AI model-behavior evaluation
Adversarial and edge-case testing across chat, multimodal inputs, tool use and indirect prompt injection.
My primary work is AI model-behavior evaluation and scientific fact-checking. I also build and operate research-information tools when a recurring verification problem needs a practical system.
Italy · EU work authorization · worldwide relocation · international contracting
Adversarial and edge-case testing across chat, multimodal inputs, tool use and indirect prompt injection.
Primary-literature review for scripts, articles and public scientific communication.
Structured records, provenance rules and maintained tools for recurring research-information problems.
Each case shows the actual public output, my contribution and the boundary of what that work demonstrates.
Across four public evaluation surfaces, I test chat, image, agent and indirect prompt-injection behavior and preserve dated evidence of the platform-reported result.
Scope: The four visible area counters sum to 109 while the profile displays 110 total breaks. Both are reported without inferring the platform’s internal aggregation. This supports evaluation and adversarial-QA applications, not penetration-testing or senior red-team engineering claims.
I contributed to 80 documented published pieces.
Paid contractor supporting an established Italian science-communication brand across evidence review, content production and website operations.
Scope: Platform metrics describe the production environment, not a personal audience. Quantified thumbnail lift is stated only when comparable analytics are available.

Yourself to Science turns scattered institutional opportunities into a public catalogue with explicit inclusion and update rules.
Scope: My contribution covers requirements, information architecture, verification, functional testing, deployment diagnosis and operations—not unaided software development.
The current product identifies MDPI references across literature-search and reference-management workflows while avoiding ambiguous title-based matches. The broader rebrand and expansion to retractions, comments and other research-integrity signals are future work, not shipped functionality.
Chrome · Edge · Firefox · Safari source · Zotero 7–9

I designed this vector diagram to show overlap among monogenic conditions associated with autism, dystonia, epilepsy and schizophrenia. The Wikimedia record exposes the source file, authorship, revision history and reuse across four Wikipedia language editions.
Open the Wikimedia source recordI work best when the objective and decision rights are explicit, feedback is direct and I can follow a problem through investigation, documentation, release and maintenance.
State what was directly observed before drawing a broader conclusion.
Document the definitions, exclusions and judgments that affect the result.
Treat updates, provenance and operational recovery as part of the work.
The targeted CVs carry the full application detail.
Independent practice · Gray Swan Proving Ground participant · Conduct self-directed adversarial testing across chat, multimodal, agentic tool-use and indirect prompt-injection settings.
Entropy for Life — Italy · Delivered 80 documented published content contributions—55 YouTube videos, 4 co-authored articles and 21 short-form pieces—for an Italian science-communication brand with 267K YouTube subscribers and 480K+ combined platform following; recurring evidence review, content production, selected thumbnails and website operations.
Yourself to Science™ · Founded and operate an open-source research-participation directory indexing more than 55 initiatives, with documented verification, provenance and metadata workflows.
The AI evaluation CV is the recommended default. The others are deliberately tailored alternatives.
Model-behavior testing, adversarial QA, evaluation operations and evidence-bound reporting.
Open CV →Scientific AI and research-data rolesScientific evidence review, research-data quality, provenance, metadata and research operations.
Open CV →Integrity and investigation rolesSource provenance, public-record research, content integrity and investigation support.
Open CV →Editorial, content-operations and audience-quality rolesScientific fact-checking, editorial coordination, audience packaging, evidence synthesis and content operations.
Open CV →Especially model-behavior evaluation, scientific fact-checking, research-data quality and knowledge-integrity work.