‹ The Index

Iris

plugin

The agent eval standard for MCP. Score every agent output for quality, safety, and cost. 12 built-in eval rules cover completeness, relevance, safety (PII detection, prompt injection), and cost thresholds. Log traces with hierarchical spans, evaluate outputs inline, and track costs across all your agents. No SDK, no code changes — add one line to your MCP config. Open source, MIT licensed.

Works with: Claude Code, Cursor, Claude Desktop, Codex CLI, Gemini CLI, Cline, Windsurf, VS Code

Category: Productivity — see all ranked ›

Work: Model evaluation · Observability

Who it is for: AI engineer · DevOps / SRE

Security audit

Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.

source ↗  ·  plugin:iris-eval/mcp-server/iris

Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›