benchmark-contamination-scan
What this skill does
Benchmark contamination scan - detects overlap between training data and 60+ evaluation datasets using n-gram and embedding similarity methods.
github/mkurman - general - 319 stars
Threat analysis
Skill info
pkg:github/mkurman/zorai@fdcd18d?skill=benchmark-contamination-scanAssessments
No risk patterns detected in this scan. A clean automated scan is a good signal, not a guarantee.
Badge
Add the Anomity scan badge for benchmark-contamination-scan to your README.
How Anomity governs this at runtime
Scan-time vetting tells you what a skill says it will do. Anomity's Endpoint Sensor sees what agents actually do: it discovers skills alongside every other AI artifact on the endpoint, and runtime governance can allow, deny, or log the tool calls a skill triggers. Policy violations route to your SIEM, Slack, email, or Jira, backed by a queryable 90-day audit trail.
Book a 30-minute demo to see your own skill inventory.
Methodology and disputes
Every skill is assessed by the Anomity Skill Intelligence engine against its public source; findings indicate risk patterns, not confirmed exploitation. Maintainer of benchmark-contamination-scan? Report an issue or request a rescan.




