Book a 30-minute demo →
Skill scan report

eval-driven-dev

View on GitHub
35 Medium Automated analysis flagged 2 potential risk patterns.

What this skill does

The skill 'eval-driven-dev' is a development tool for improving AI applications through evaluation-driven development. It focuses on setting up QA, adding tests, and evaluating Python-based LLM applic

github/github - Potential Code Injection - 37k stars

ThreatsPotential Code Injection Insecure Package Installation

Threat analysis

Potential Code Injection1 finding
Insecure Package Installation1 finding

Skill info

Namegithub/eval-driven-dev
Registrygithub
Versiondae77f2
PURLpkg:github/github/awesome-copilot@dae77f2?skill=eval-driven-dev
Stars37k

Assessments (2)

Potential Code Injection1 finding HIGH
HIGH

Potential Code Injection via local-llm-review

resources/setup.sh

The script uses `uv add` and `uv pip install` to install packages, but the command is cut off and incomplete. This could be a red flag if the script is attempting to install untrusted or malicious pac
Insecure Package Installation1 finding MEDIUM
MEDIUM

Insecure Package Installation via local-llm-review

resources/setup.sh

The script attempts to install 'pixie-qa[all]' without explicitly specifying a trusted source or version pinning. This could lead to the installation of potentially insecure or malicious versions of t

Badge

Add the Anomity scan badge for eval-driven-dev to your README.

Anomity Skill Check badge

Markdown
[![Anomity Skill Check](https://anomity.ai/skills/badge.svg)](https://anomity.ai/skills/github/github/eval-driven-dev/)
HTML
<a href="https://anomity.ai/skills/github/github/eval-driven-dev/"><img src="https://anomity.ai/skills/badge.svg" alt="Anomity Skill Check"></a>
Image URL
https://anomity.ai/skills/badge.svg

How Anomity governs this at runtime

Scan-time vetting tells you what a skill says it will do. Anomity's Endpoint Sensor sees what agents actually do: it discovers skills alongside every other AI artifact on the endpoint, and runtime governance can allow, deny, or log the tool calls a skill triggers. Policy violations route to your SIEM, Slack, email, or Jira, backed by a queryable 90-day audit trail.

Book a 30-minute demo to see your own skill inventory.

Methodology and disputes

Every skill is assessed by the Anomity Skill Intelligence engine against its public source; findings indicate risk patterns, not confirmed exploitation. Maintainer of eval-driven-dev? Report an issue or request a rescan.

Ask AI about Anomity
ChatGPT Claude Perplexity Google AI Grok