‹ The Index

Memoriant Eval Sandbox Skill

plugin

AI agent evaluation sandbox for Claude Code. Holdout scenario testing with Doer/Judge/Adversary/Observer roles, probabilistic satisfaction scoring, and append-only JSONL audit trails with integrity hashes. Test your agents before deploying them. Includes full Python evaluation framework.

Works with: Claude Code, Cursor, Claude Desktop, Codex CLI, Gemini CLI, Cline, Windsurf, VS Code

Category: Dev Tools & CI — see all ranked ›

Security audit

Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.

source ↗  ·  plugin:nathanmaine/memoriant-eval-sandbox-skill/memoriant-eval-sandbox-skill

Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›