‹ The Index
Evals
remote
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Works with: Claude Code, Cursor, Claude Desktop, Codex CLI, Gemini CLI, ChatGPT
Category: AI & Agents — see all ranked ›
- tashan score: 34.0
- Health: active
- GitHub stars: 1
- Contributors: 3
- License: NOASSERTION
Security audit
Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.
source ↗ · registry:com.completionkit/evals
Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›