Memoriant Eval Sandbox Skill
AI agent evaluation sandbox for Claude Code. Holdout scenario testing with Doer/Judge/Adversary/Observer roles, probabilistic satisfaction scoring, and append-only JSONL audit trails with integrity hashes. Test your agents before deploying them. Includes full Python evaluation framework.
Works with: Claude Code, Cursor, Claude Desktop, Codex CLI, Gemini CLI, Cline, Windsurf, VS Code
Category: Dev Tools & CI — see all ranked ›
- tashan score: 25.0
- Adoption: 1 repos
- Health: active
- GitHub stars: 0
- Contributors: 1
- License: MIT
Security audit
Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.
source ↗ · plugin:nathanmaine/memoriant-eval-sandbox-skill/memoriant-eval-sandbox-skill
Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›