# Evalview

> Behavior regression testing for AI agents. EvalView detects when your agent's behavior drifts e.g changed tool calls, different outputs, or degraded quality even when traditional tests still pass. Snapshot baselines, check for regressions, and auto-heal flaky results, all from inside Claude Code.

## Facts
- Page: https://tashan.sh/capability/plugin-hidai25-eval-view-evalview
- tashan id: plugin:hidai25/eval-view/evalview
- Source: https://github.com/hidai25/eval-view
- Type: plugin
- Category: devtools
- tashan score: 62.0 / 100
- Adoption: 34.0
- Upkeep: 93.0
- Freshness: 84.0
- Evidence coverage: 84% of the inputs this score can use
- Health: active
- Instruction depth: solid
- GitHub stars: 124
- License: Apache-2.0
- Official: no

## Install

```sh
/plugin marketplace add anthropics/claude-plugins-community
/plugin install evalview@claude-community
```

## Security audit
Not scanned. We audit npm-published capabilities; this one has no npm package we can resolve, or has not reached the queue. This is not a clean bill of health.

---
Measured 2026-09-13 by tashan (https://tashan.sh) from public evidence. Scorer s5.
