# Promptfoo Evals

> Teaches AI coding agents to create and maintain promptfoo eval suites. Encodes best practices like deterministic assertions first, file-based test organization, correct environment variable syntax, and proper hallucination/faithfulness checks. Includes a cheatsheet covering 30+ assertion types, provider patterns for OpenAI/Anthropic/Google/Bedrock/HTTP/Python/JS, and test patterns for JSON validation, dataset-driven scaling, and CI-friendly runs.

## Facts
- Page: https://tashan.sh/capability/plugin-promptfoo-promptfoo-promptfoo-evals
- tashan id: plugin:promptfoo/promptfoo/promptfoo-evals
- Source: https://github.com/promptfoo/promptfoo
- Type: plugin
- Category: devtools
- tashan score: 76.0 / 100
- Adoption: 63.0
- Upkeep: 93.0
- Freshness: 84.0
- Evidence coverage: 84% of the inputs this score can use
- Health: active
- Instruction depth: not yet graded
- GitHub stars: 23,659
- License: MIT
- Official: no

## Install

```sh
/plugin marketplace add anthropics/claude-plugins-community
/plugin install promptfoo-evals@claude-community
```

## Security audit
Not scanned. We audit npm-published capabilities; this one has no npm package we can resolve, or has not reached the queue. This is not a clean bill of health.

---
Measured 2026-09-13 by tashan (https://tashan.sh) from public evidence. Scorer s5.
