deepeval
What this skill does
DeepEval is an evaluation framework for AI agents and LLM applications. It enables users to instrument their applications, curate datasets, run evaluations, and iterate on failures to improve AI agent
github/confident-ai - general - 17.5k stars
Threat analysis
Skill info
pkg:github/confident-ai/deepeval@f979644?skill=deepevalAssessments
No risk patterns detected in this scan. A clean automated scan is a good signal, not a guarantee.
Badge
Add the Anomity scan badge for deepeval to your README.
How Anomity governs this at runtime
Scan-time vetting tells you what a skill says it will do. Anomity's Endpoint Sensor sees what agents actually do: it discovers skills alongside every other AI artifact on the endpoint, and runtime governance can allow, deny, or log the tool calls a skill triggers. Policy violations route to your SIEM, Slack, email, or Jira, backed by a queryable 90-day audit trail.
Book a 30-minute demo to see your own skill inventory.
Methodology and disputes
Every skill is assessed by the Anomity Skill Intelligence engine against its public source; findings indicate risk patterns, not confirmed exploitation. Maintainer of deepeval? Report an issue or request a rescan.




