Databricks Synthetic Data Gen
Generate realistic synthetic data using Spark + Faker (strongly recommended). Supports serverless execution, multiple output formats (Parquet/JSON/CSV/Delta), and scales from thousands to millions of rows. For small datasets (<10K rows), can optionally generate locally and upload to volumes. Use when user mentions 'synthetic data', 'test data', 'generate data', 'demo dataset', 'Faker', or 'sample data'.
Works with: Claude Code (native) · Cursor, Codex CLI (manual)
native: this artifact type is that client's own format
Category: Other — see all ranked ›
Install (Claude Code):
cp -r databricks-synthetic-data-gen ~/.claude/skills/- Adoption: 1 repos
Security audit
Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.
source ↗ · skill:databricks/databricks-synthetic-data-gen
Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›
Measured 2026-08-03 · scorer s5 · how · something wrong here?