# Docx Smart Extractor

> Extract and analyze Word documents (1MB-50MB+) with minimal token usage. Lossless extraction of all text, tables, formatting, and document structure while achieving 10-50x token reduction through local extraction, semantic chunking by headings, and intelligent caching. Perfect for policy documents, contracts, technical reports, and any large Word document.

## Facts
- Page: https://tashan.sh/capability/plugin-diegocconsolini-claudeskillcollection-docx-smart-extractor
- tashan id: plugin:diegocconsolini/claudeskillcollection/docx-smart-extractor
- Source: https://github.com/diegocconsolini/ClaudeSkillCollection
- Type: plugin
- Category: security
- tashan score: 33.0 / 100
- Adoption: 7.0
- Upkeep: 59.0
- Freshness: 88.0
- Evidence coverage: 84% of the inputs this score can use
- Health: active
- Instruction depth: not yet graded
- License: MIT
- Official: no

## Install

```sh
/plugin marketplace add diegocconsolini/ClaudeSkillCollection
/plugin install docx-smart-extractor@security-compliance-marketplace
```

## Security audit
Not scanned. We audit npm-published capabilities; this one has no npm package we can resolve, or has not reached the queue. This is not a clean bill of health.

---
Measured 2026-08-05 by tashan (https://tashan.sh) from public evidence. Scorer s5.
