Now in early access, book a 30-minute demo →
Skill scan report

benchmark-contamination-scan

View on GitHub
0 Clean No risk patterns detected.

What this skill does

Benchmark contamination scan - detects overlap between training data and 60+ evaluation datasets using n-gram and embedding similarity methods.

github/mkurman - general - 319 stars

ThreatsNo threats detected

Threat analysis

No risk categories triggeredclean

Skill info

Namemkurman/benchmark-contamination-scan
Registrygithub
Versionfdcd18d
PURLpkg:github/mkurman/zorai@fdcd18d?skill=benchmark-contamination-scan
Stars319

Assessments

No risk patterns detected in this scan. A clean automated scan is a good signal, not a guarantee.

Badge

Add the Anomity scan badge for benchmark-contamination-scan to your README.

Anomity Skill Check badge

Markdown
[![Anomity Skill Check](https://anomity.ai/skills/badge.svg)](https://anomity.ai/skills/github/mkurman/benchmark-contamination-scan/)
HTML
<a href="https://anomity.ai/skills/github/mkurman/benchmark-contamination-scan/"><img src="https://anomity.ai/skills/badge.svg" alt="Anomity Skill Check"></a>
Image URL
https://anomity.ai/skills/badge.svg

How Anomity governs this at runtime

Scan-time vetting tells you what a skill says it will do. Anomity's Endpoint Sensor sees what agents actually do: it discovers skills alongside every other AI artifact on the endpoint, and runtime governance can allow, deny, or log the tool calls a skill triggers. Policy violations route to your SIEM, Slack, email, or Jira, backed by a queryable 90-day audit trail.

Book a 30-minute demo to see your own skill inventory.

Methodology and disputes

Every skill is assessed by the Anomity Skill Intelligence engine against its public source; findings indicate risk patterns, not confirmed exploitation. Maintainer of benchmark-contamination-scan? Report an issue or request a rescan.

Ask AI about Anomity
ChatGPT Claude Perplexity Google AI Grok