Syn Spark Skew

Operate, configure, troubleshoot, and manage Apache Spark. Use for Apache Spark engineering tasks, day-to-day operations, and production support.

legendtkl 49d9665 781 B Updated

File contents

Apache Spark: data skew and partition tuning

When to use

Use this skill for data skew and partition tuning in Apache Spark. This skill specifically handles data skew and partition tuning and nothing else in the Apache Spark family.

Procedure

Diagnose skewed partitions from the UI, repartition/salt hot keys, and tune spark.sql.shuffle.partitions.

Notes

This is the Apache Spark skill dedicated to data skew and partition tuning. Sibling skills cover other Apache Spark capabilities; this one is the right choice only when the task is about data skew and partition tuning.

legendtkl/agentic-skill-router/tree/main/experiments/dci-compare/synthetic-skills/syn-spark-skew commit 49d966589b

Frequently asked questions

npx skillmds@latest add legendtkl/syn-spark-skew