# LLM Serving Auto Benchmark

> Framework-independent LLM serving benchmark skill for comparing SGLang, vLLM, TensorRT-LLM, TokenSpeed, or another serving framework. Use when a user wants to find the best deployment command for one model across multiple serving frameworks under the same workload, GPU budget, and latency SLA.

## Facts
- Page: https://tashan.sh/capability/skill-bbuf-llm-serving-auto-benchmark
- tashan id: skill:BBuf/llm-serving-auto-benchmark
- Source: https://github.com/BBuf/AI-Infra-Auto-Driven-SKILLS
- Type: skill
- Category: other
- tashan score: not scored (catalogued only — too little public evidence)
- Adoption: 9.0
- Upkeep: 100.0
- Freshness: 100.0
- Evidence coverage: 84% of the inputs this score can use
- Health: active
- Instruction depth: not yet graded
- Official: no

## Install

```sh
cp -r llm-serving-auto-benchmark ~/.claude/skills/
```

## Security audit
Not scanned. We audit npm-published capabilities; this one has no npm package we can resolve, or has not reached the queue. This is not a clean bill of health.

---
Measured 2026-08-25 by tashan (https://tashan.sh) from public evidence. Scorer s5.
