AI model-behavior evaluation
Adversarial and edge-case testing across chat, multimodal inputs, tool use and indirect prompt injection.
I retrieve and verify information, investigate data and evidence quality, and test model behavior. When the work needs a durable system, I define the research process, test the result and keep it maintained.
Italy · EU/EEA work-authorised · sponsored international relocation · B2B engagements
Adversarial and edge-case testing across chat, multimodal inputs, tool use and indirect prompt injection.
Finding primary sources, tracing claims and verifying scientific evidence for public-facing work.
Provenance, metadata, source reconciliation, validation, public-source investigation and information-quality review.
Structured records, provenance rules, maintained tools and data visualizations for recurring research-information problems.
Each case shows the actual public output, my contribution and the boundary of what that work demonstrates.

I tested chat, image, agent and indirect prompt-injection behavior and preserved dated public evidence that keeps Proving Ground and Arena metrics separate.
Scope: The four visible area counters sum to 112 while the profile displays 113 total breaks. Both are reported without inferring the platform's internal aggregation, and the record is dated evaluation evidence rather than a model-wide conclusion.
Public work record across evidence review, content production and publishing.
Paid contractor supporting an established Italian science-communication brand through evidence review, English-to-Italian scientific localization, content production and WordPress website management.
Scope: Platform metrics describe the production environment, not a personal audience. Quantified thumbnail lift is stated only when comparable analytics are available.
My public record includes consumer-genomics privacy research, corporate-source reconciliation, archival recovery, content-governance review and biomedical evidence synthesis—not a generic claim of OSINT familiarity.
Scope: The work represents public-source research and collaborative knowledge governance. It does not establish company liability, personal misconduct, editor affiliation, an independent legal judgment or original clinical research.

Yourself to Science turns scattered institutional opportunities into a public catalogue with explicit inclusion and update rules.
Scope: My technical responsibilities cover requirements, information architecture, verification, functional testing, deployment diagnosis and ongoing maintenance; implementation is AI-assisted.
Notandia identifies articles from scrutinized publishers such as MDPI and Frontiers, checks Crossref/Retraction Watch for formal notices, and adds precise MDPI reference detection in Zotero. Publisher context is not an article-quality verdict.
Chrome · Edge · Firefox · Safari source · Zotero 7–9

I designed this vector diagram to show overlap among monogenic conditions associated with autism, dystonia, epilepsy and schizophrenia. The Wikimedia record exposes the source file, authorship, revision history and reuse across four Wikipedia language editions.
Open the Wikimedia source recordI work best when the objective and decision rights are explicit, feedback is direct and I can follow a problem through investigation, documentation, release and maintenance.
State what was directly observed before drawing a broader conclusion.
Document the definitions, exclusions and judgments that affect the result.
Treat updates, provenance and recovery from problems as part of the work.
The targeted CVs carry the full application detail.
Independent practice · Gray Swan Proving Ground participant · Conduct self-directed adversarial testing across chat, multimodal, agentic tool-use and indirect prompt-injection settings.
Entropy for Life — Italy · Delivered 80 documented published content contributions for an Italian science-communication brand with 267K YouTube subscribers and 480K+ combined platform following; recurring evidence review, selected visual packaging and WordPress website management.
Yourself to Science™ · Founded and operate an open-source research-participation directory indexing more than 55 initiatives, with documented verification, provenance and metadata workflows.
Use the data-quality and research-analysis CV as the general default; the others are focused specialist documents.
Information retrieval, data quality, source verification, data visualization and research support.
Open CV →Specialist CV for AI evaluation and adversarial-testing rolesModel testing, adversarial QA, test planning and evidence reporting.
Open CV →Integrity, trust & safety and investigation rolesConsumer-genomics privacy, archival OSINT, source-quality review, public-record research, content governance and investigation support.
Open CV →Editorial, community-coordination and content-quality rolesScientific fact-checking, editorial coordination, research participation, community engagement and content production.
Open CV →Especially research analysis, scientific verification, model-behavior evaluation and knowledge-integrity work.