‹ The Index

Databricks Synthetic Data Gen

skill

Generate realistic synthetic data using Spark + Faker (strongly recommended). Supports serverless execution, multiple output formats (Parquet/JSON/CSV/Delta), and scales from thousands to millions of rows. For small datasets (<10K rows), can optionally generate locally and upload to volumes. Use when user mentions 'synthetic data', 'test data', 'generate data', 'demo dataset', 'Faker', or 'sample data'.

Works with: Claude Code (native)  ·  Cursor, Codex CLI (manual)
native: this artifact type is that client's own format

Category: Other — see all ranked ›

Install (Claude Code):

cp -r databricks-synthetic-data-gen ~/.claude/skills/

Security audit

Not scanned yet. We audit npm-published capabilities for known advisories, install-time scripts and permission surface; this one has no npm package we can resolve, or has not reached the queue.

source ↗  ·  skill:databricks/databricks-synthetic-data-gen

Already running this? npx tashan-cli doctor checks your whole config against the Index — how it works ›

Measured 2026-08-03  ·  scorer s5  ·  how  ·  something wrong here?