GCP Spark

Develops and executes Spark code on Managed Spark on Google Cloud (Dataproc Clusters and Serverless). Reads and writes data using BigLake Iceberg catalogs, BigQuery and Spanner. Debugs execution failures. Use when: - Writing Spark ETL pipelines on Google Cloud Platform. - Training or running inference with Machine Learning models with spark on Google Cloud Platform. - Managing Spark clusters, jobs, batches, and interactive sessions. Don't use when: - Writing generic Python scripts that don't use Spark. - Performing simple SQL queries that can be done directly in BigQuery.

gemini-cli-extensions Updated

File contents

gemini-cli-extensions/data-agent-kit-starter-pack/tree/main/skills/gcp-spark commit eab7ca0736

Frequently asked questions

npx skillmds@latest add gemini-cli-extensions-data-agent-kit-starter-pack/gcp-spark