Spark

Use when building distributed data pipelines with Apache Spark. Covers partitioning, shuffles, skew, joins, caching, and reading the Spark UI to find why a job is slow.

nimadorostkar 7a3a13e 5.4 KB Updated

File contents

nimadorostkar/Claude-Skills-collection/tree/main/skills/data/spark commit 7a3a13e4f7

Frequently asked questions

npx skillmds@latest add nimadorostkar/spark