# Mario Marcolongo — Data Quality & Information Retrieval | AI Evaluation & Adversarial Testing | Scientific Fact-Checking & Evidence Synthesis (AI / LLM Compatible Markdown)
> Information Retrieval · Evidence Synthesis · AI Evaluation
> I make information and AI systems more reliable.
> Work Authorization: EU/EEA work-authorised; Switzerland: EU/EFTA employment route once an offer is secured; open to employer-sponsored work authorisation elsewhere. International B2B engagements are available where the contracting arrangement is compliant.
> Email: me@mariomarcolongo.com | Web: https://mariomarcolongo.com | ORCID: 0000-0003-2846-7115

---

## Executive Summary
My work spans information retrieval, data quality, scientific fact-checking, AI evaluation and research operations. It includes eight years of auditable public-source and structured-data work, a maintained research-participation directory and self-directed model-behavior testing. On 29 July 2026, the Gray Swan Proving Ground profile displayed rank #74 (top 6%) and 113 total breaks. Entropy for Life work covers 80 documented published content contributions: 55 YouTube videos · 4 articles · 21 short-form pieces. Technical delivery is AI-assisted, with personal responsibility for requirements, verification, deployment and maintenance. Current open-source product work includes Notandia (formerly MDPI Filter), a continuing browser and Zotero research-integrity tool.

> **My technical work uses AI-assisted implementation: I define requirements, inspect implementation behavior, test releases, diagnose functional problems, guide iterations, deploy releases and maintain services.**

---

## Core Competencies & Technical Skills
- Information Retrieval & Verification: Primary-source and bibliographic search, claim tracing, source discovery, query refinement, source-quality assessment and cross-source corroboration.
- Data Quality & Provenance: Structured metadata, entity reconciliation, validation rules, documentation, taxonomy design and public-record verification.
- AI Evaluation & Adversarial Testing: Exploratory model-behavior testing, prompt and jailbreak analysis, multi-turn behavior, multimodal inputs, agentic tool-use and indirect prompt injection.
- Scientific Verification & Evidence Synthesis: Primary-source fact-checking, bibliographic research, claim decomposition, evidence screening, source-quality assessment and cross-source corroboration.
- Data Visualization: SVG vector diagrams, Tableau Public, Flourish and open-data publication.
- Evaluation Planning & Reporting: Test planning, evidence capture, reproducibility notes, taxonomy thinking, issue classification, reporting that separates evidence from inference and mitigation-retesting concepts.
- Technical Delivery: Requirements definition, codebase reading and behavior inspection, functional testing, issue diagnosis, deployment and service maintenance.
- AI-Assisted Implementation: Uses coding agents for implementation support while personally defining requirements, inspecting structure and behavior, testing results, guiding revisions, deploying releases and maintaining services.
- Open Science & Structured Data: Wikimedia, Wikidata, FAIRsharing, Zenodo, ENA records, JSON-LD, RDF Turtle/VoID, OpenAPI and MCP interfaces.
- Web & Cloud Delivery: WordPress, HTML/CSS modification, Git/GitHub, JSON, REST APIs, AWS Lambda, Cloudflare Pages, DNS/SSL and CI/CD maintenance.
- Languages: Italian — native. English — C1 overall (EF SET 68/100), with advanced technical reading and professional/technical writing.

---

## Professional Experience

### Founder & Project Lead — Yourself to Science™
*Aug 2024 — Present* · **[Independent project]**

- Founded and operate an open-source research-participation directory indexing 55 resources; as of 27 July 2026, 37 unique Wikidata items use yourselftoscience.org as a reference URL (P854).
- Defined the inclusion model, verification workflow, provenance fields, licensing structure and machine-readable metadata requirements.
- Use AI-assisted implementation while personally defining requirements, inspecting code structure and behavior, testing releases and maintaining the service.

### Independent AI Evaluator — Independent practice · Gray Swan Proving Ground participant
*Jul 2026 — Present* · **[Independent practice]**

- Conduct self-directed testing of LLM instruction handling, policy boundaries and edge cases across chat, image, agentic tool-use and indirect prompt-injection settings.
- Reached #74 on the Proving Ground (top 6%) with 113 platform-displayed total breaks on 29 July 2026; the same profile displayed Arena rank #365, 28 global unique breaks, 1,120 points and 255 submissions.
- Document the visible 112/113 discrepancy and separate platform-reported outcomes from independent verification, security certification or model-wide conclusions.

### Science Writer & Fact-Checker / Website Manager (WordPress) — Entropy for Life — Italy
*Jun 2023 — Present* · **[Independent contractor]**

- Delivered 80 documented published content contributions: 55 YouTube videos, 4 co-authored articles and 21 short-form pieces.
- Own recurring primary-literature research and scientific fact-checking; depending on the assignment, translate evidence into scripts, data analyses, visualizations, presentation slides, on-screen assets and short-form content.
- Develop selected thumbnail concepts and visual packaging independently or with the video editor, using click-through rate, watch time, retention and immediate attention capture as explicit design criteria.
- Designed and built entropyforlife.it in WordPress and manage responsive design, publishing, OVHcloud hosting, DNS, SSL and technical SEO; formally acknowledged in the Mondadori book Italiani veri for scientific-literature research and error detection.

### Volunteer Research Assistant & Focus-Group Co-Facilitator — Department of Developmental Psychology and Socialisation (DPSS), University of Padua
*Nov 2022 — 2025* · **[Volunteer research collaboration supervised by Marta Panzeri]**

- Served as lead or co-facilitator across approximately 4–5 recorded Zoom focus-group sessions, typically lasting 1–2 hours, with autistic participants discussing sensitive sexuality and relationship topics.
- Co-developed the protocol, including pseudonymous naming, explicit recorded consent, optional captions and written-chat participation, timed turn-taking, scripted prompts, recording boundaries and two-person facilitation handoffs.
- Supported participant recruitment, bibliographic research and technical preparation; contributed to structuring participant experience and accessibility considerations and coordinated with Marta Panzeri, researchers and a second volunteer facilitator.
- Public attribution to Marta Panzeri and the Department of Developmental Psychology and Socialisation is included with permission; no participant information or confidential session content is disclosed.

### Scientific Contributor & Structured-Data Editor — Wikipedia, Wikidata & Wikimedia Commons
*Mar 2018 — Present* · **[Public contribution record]**

- Completed 4,317 auditable contributions across Wikimedia projects as of July 2026.
- Check citations, reconcile conflicting sources, improve structured metadata and create scientific visualizations adopted across four Wikipedia language editions.

---

## Research & Open Science Infrastructure

### Personal Genomics Workflow & Open-Data Record (41×) — European Nucleotide Archive — PRJEB109744 / SAMEA121950568
*Jan 2026* · **[Personal open-data project]** | Website: https://www.ebi.ac.uk/ena/browser/view/PRJEB109744 | GitHub: https://github.com/jnton/git-nome

- Donated personal 41× whole-genome sequencing raw paired-end FASTQ reads to the public domain under ENA BioSample SAMEA121950568.
- Defined analytical requirements and used Terra.bio cloud workflows to produce GRCh38 BAM alignments and VCF variant-call files.
- Specified requirements for and iteratively validated a downstream workflow using Plink2 and Nextflow pgsc_calc with ancestry projection.
- Used AI-assisted Python scripts for VEP-annotated VCF extraction, multi-trait polygenic-score summaries and variant filtering; inspected outputs and iterated requirements rather than independently developing the code.

### Scientific Data Visualization & Evidence Synthesis — Wikimedia Commons, Tableau Public & Flourish
*2023 — Present* · **[Independent project]** | Tableau: https://public.tableau.com/app/profile/mario.marcolongo/vizzes | Flourish: https://app.flourish.studio/@Digressivo

- Published more than 70 empirical public-health, epidemiological and biomedical visualizations across Wikimedia Commons, Tableau Public and Flourish.
- Created an original vector Euler diagram of overlapping monogenic clinical phenotypes adopted across four Wikipedia language editions.
- Synthesized primary datasets into open-access charts, maps and vector evidence records.

---

## Deployed Systems & Open-Source Projects

### Model Behavior & Adversarial Evaluation
**Independent AI Evaluator**

Self-directed model-behavior evaluation conducted through the Gray Swan Proving Ground. The 29 July 2026 profile snapshot displays rank #74, top 6%, and 113 total breaks. The same snapshot shows 255 Arena submissions, 28 global unique breaks, 1,120 points and Arena rank #365. Aggregate counts are presented with explicit evidence limitations; complete prompts, outputs, model versions and adjudication materials are not reproduced.

### Yourself to Science™ | Open Research-Participation Directory
**Founder & Project Lead** | Website: https://yourselftoscience.org | GitHub: https://github.com/yourselftoscience/yourselftoscience.org | DOI: https://doi.org/10.5281/zenodo.15109359 | FAIRsharing: https://doi.org/10.25504/FAIRsharing.d3d487

Founded, designed and operate an open-source research-participation directory indexing 55 resources. As of 27 July 2026, 37 unique Wikidata items use yourselftoscience.org as a reference URL (P854). Defined the inclusion model, verification workflow, provenance fields, licensing structure and machine-readable metadata requirements, including JSON-LD, RDF Turtle/VoID, OpenAPI and an MCP interface. Technical implementation is AI-assisted and personally verified through requirements, code reading, functional testing and maintenance.

### Entropy for Life — Science Writing, Fact-Checking & Website Management
**Science Writer & Fact-Checker / Website Manager (WordPress)**

Conduct recurring primary-literature research and scientific fact-checking across 80 documented published content contributions: 55 YouTube videos, 4 co-authored articles and 21 short-form pieces. Entropy for Life is an Italian science-communication brand with 267K YouTube subscribers, 36,524,137 channel views, 159K Instagram followers and 54K TikTok followers as of 26 July 2026. Depending on the assignment, also translate evidence into scripts, data analyses, visualizations, slides, on-screen assets, short-form content and selected thumbnail concepts or production. Designed and built entropyforlife.it in WordPress and operate its responsive design, publishing and OVHcloud technical stack. The audience belongs to the brand, and cross-platform totals are not counts of unique people.

### Notandia — formerly MDPI Filter | Browser Extension & Zotero Plugin
**Creator & Product Lead** | Website: https://mariomarcolongo.com/notandia | Chrome Store: https://chromewebstore.google.com/detail/mdpi-filter/comknkeimaaadpiopddjoknflbmjeccp | Edge Store: https://microsoftedge.microsoft.com/addons/detail/mdpi-filter/efonlkldplkaeekpiajloajjmkappjgi | GitHub: https://github.com/notandia/browser-extension

Created and maintain Notandia, the public continuation and expansion of MDPI Filter. The browser extension helps users identify articles from publishers whose editorial and peer-review practices have attracted scrutiny—including MDPI and Frontiers—and choose how matching articles should appear, including context, badges, highlights, dimming or hiding. It also uses Crossref/Retraction Watch data to check for formal notices such as retractions, corrections, expressions of concern, withdrawals, duplicate-publication findings and reinstatements. I define product requirements and evidence hierarchies, test cross-browser and Zotero behavior, inspect API and implementation logic, reproduce failures, guide AI-assisted changes, and manage documentation, release verification and deployment. The Zotero plugin currently focuses on precise MDPI item and reference detection, including structured PubMed Central evidence and exact citation highlighting. Publisher-level context is controlled by the user and does not treat every journal or article as equivalent. Existing store identities and compatibility-sensitive identifiers are retained so the same product lineage can continue receiving updates. Current source repositories: https://github.com/notandia/browser-extension and https://github.com/notandia/zotero-plugin.

### English Wikipedia Link Converter | Telegram Bot
**Creator & Project Lead** | GitHub: https://github.com/jnton/english-wikipedia-link-converter-telegram-bot | Bot: https://t.me/ToEnWikipediaBot

Specified the behavior and deployment requirements for a Telegram bot running on AWS Lambda and API Gateway. Used AI-assisted implementation, tested private, group and inline workflows, diagnosed deployment problems and maintained GitHub Actions releases.

### Emergent Humanity | Interactive Network Narrative
**Creator & Project Lead** | Website: https://jnton.github.io/emergent-humanity/ | GitHub: https://github.com/jnton/emergent-humanity

Developed the concept, narrative structure, interaction requirements and behavioral specifications for an evolving browser-based network simulation. Implementation was produced through AI-assisted workflows and personally tested and iterated.

---

## Education & Certifications

- **Information Technology — Open UAS Path Studies (120 ECTS)** (Aug 2026 — Present) — Metropolia University of Applied Sciences · Bachelor's-level ICT path studies · After completing 120 ECTS, eligible to apply to the BEng IT programme (admission not yet granted) · Curriculum: software development, SQL and relational databases, Unix/Linux, data structures and algorithms, CCNA networking, cloud computing, cybersecurity, ethical hacking, engineering mathematics and applied AI · Programme: https://www.metropolia.fi/en/study-at-metropolia/open-university/path-studies/it-online
- **EF SET English Certificate 68/100 (C1 overall)** (Mar 2024) — EF Standard English Test · Credential: https://cert.efset.org/en/eJz39v
- **Career Essentials in Generative AI** (Mar 2024) — Microsoft and LinkedIn · Credential: https://www.linkedin.com/learning/certificates/c4f1f59578e3ac2567787e262e3b2ec55debf96bbbced4d34d5edaa821d9e6d9
- **GALENOS Crowd Evidence Synthesis Training** (May 2026) — Cochrane Crowd & GALENOS · Systematic-review screening training

## Contact & Identifiers
- [Email](mailto:me@mariomarcolongo.com): Direct email contact
- [Website](https://mariomarcolongo.com): Main portfolio homepage
- [ORCID](https://orcid.org/0000-0003-2846-7115): ORCID scientific researcher profile
- [LinkedIn](https://www.linkedin.com/in/mario-marcolongo): Professional LinkedIn profile
- [GitHub](https://github.com/jnton): GitHub profile and open-source repositories
- [Agent-Readiness Audit](https://isitagentready.com/mariomarcolongo.com?checks=robotsTxt%2Csitemap%2ClinkHeaders%2CdnsAid%2CmarkdownNegotiation%2CrobotsTxtAiRules%2CcontentSignals%2CwebBotAuth%2CapiCatalog%2CoauthDiscovery%2CoauthProtectedResource%2CauthMd%2CmcpServerCard%2Ca2aAgentCard%2CagentSkills%2CwebMcp%2Cx402%2Cmpp%2Cucp%2Cacp): Agent-discovery and machine-readable endpoints implemented; live audit to be rerun after deployment.
- [A2A Agent Card](https://mariomarcolongo.com/.well-known/agent-card.json): Agent-to-Agent discovery card
- [Genomic Pipeline Codebase](https://github.com/jnton/git-nome): Personal genomics workflow and open-data pipeline
- [Portfolio SSOT Codebase](https://github.com/jnton/mariomarcolongo): Source code and Single Source of Truth for this portfolio

