Data & Analytics Agent Skills

Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.

Data & Analytics

1,725 skills
agentskillexchange
mongodb-mcp-server
Installs the official MongoDB Node.js driver and verifies package integrity for building MongoDB-backed applications.
28
redpanda-data
connect-debugging
Diagnoses and validates Redpanda Connect pipelines using linting, dry-run connection tests, logging, metrics, tracing, and health endpoints, including enterprise feature troubleshooting.
6 · bundle
redpanda-data
connect-cdc-salesforce
Streams Salesforce change data capture and platform events into Redpanda or Kafka using Redpanda Connect's salesforce_cdc input, covering setup, configuration, and operational details.
6 · bundle
pranavnagrecha
fhir-data-mapping
Maps FHIR R4 clinical resources (Patient, Observation, Condition, CarePlan, CodeableConcept) to Salesforce Health Cloud objects, including prerequisite configuration and cardinality handling.
15 · bundle
comeonoliver
weather
Fetches current weather and forecasts for any location using free, keyless APIs like wttr.in and Open-Meteo.
61
comeonoliver
gwas-prs
Calculates polygenic risk scores from 23andMe or AncestryDNA genotype files using PGS Catalog scoring files, then estimates population percentiles and risk categories.
61
dvcrn
whoop
Fetches your latest WHOOP recovery, sleep, and strain data and generates a short set of daily suggestions.
32 · bundle
lingxling
vaex
Process and analyze tabular datasets larger than RAM using lazy, out-of-core DataFrames, with fast aggregations, visualization, and machine learning integration.
253 · bundle
lingxling
rdkit
Provides guidance for using RDKit to read and write molecular structures, calculate descriptors, generate fingerprints, perform substructure searches, and handle chemical reactions.
253 · bundle
lingxling
flowio
Parse FCS (Flow Cytometry Standard) files v2.0-3.1, extracting events as NumPy arrays, metadata, and channel information, with support for creating and modifying FCS files.
253 · bundle
lingxling
polars
High-performance DataFrame library for Python ETL, analytics, and pandas migration. Use for expression-based data manipulation with lazy query optimization, parallel execution, streaming out-of-core processing, Arrow interoperability, and optional GPU execution.
253 · bundle
lingxling
onekgpd
Queries the 1000 Genomes Project dataset (3,202 whole-genome-sequenced individuals, GRCh38) at the level of individual participants, returning variants, carriers, and relatedness with allele frequencies and annotations.
253 · bundle
lingxling
convex
Design schemas, write TypeScript functions, and manage real-time subscriptions, auth, file storage, scheduling, and deployment for Convex backends.
253
mocchalera
review-roughcut
Reviews a video rough cut by generating deterministic metrics, critiquing the edit against brief and blueprint, and producing review_report.yaml and review_patch.json artifacts.
3 · bundle
leandrobenjaminl
db-admin
Administra bases de datos PostgreSQL, MySQL, Redis y SQLite: diseña esquemas, optimiza queries, configura migraciones, replicación y backups.
0
leandrobenjaminl
data-design
Define el enfoque, las herramientas y el pipeline de análisis antes de escribir código, eligiendo entre SQL, Python o un enfoque híbrido según la pregunta y los datos.
0
leandrobenjaminl
file-formats
Lee y escribe datos en múltiples formatos con Pandas — CSV, Excel, Parquet, JSON, Feather — y elige el formato óptimo según el caso.
0
leandrobenjaminl
data-analysis
Analiza datasets con Pandas y NumPy: explora distribuciones, correlaciones y patrones, y aplica tests de hipótesis para extraer conocimiento no obvio.
0 · bundle
leandrobenjaminl
shared-git-data
Sets up Git-based version control for data science projects, handling notebooks, datasets, and pipelines with DVC and nbstripout.
0
leandrobenjaminl
database-connections
Connect to PostgreSQL, MySQL, SQLite, and SQL Server databases using SQLAlchemy, Pandas, and DuckDB. Read and write tables, manage sessions, and handle connection strings securely.
0
leandrobenjaminl
notebook-integration
Guides creating clean, reproducible Jupyter notebooks with structured workflows, best practices, and IPython magic for analysis and presentation.
0 · bundle
martc03
gov-contracts
Search SAM.gov for contract opportunities and registered entities, and query USASpending.gov for federal award data.
5
martc03
gov-public-health
Queries CDC open data and WHO Global Health Observatory indicators for public health intelligence, covering disease surveillance, vaccinations, and mortality.
5
martc03
gov-competitive-intel
Gather competitive intelligence on companies by searching SEC filings, recent news, federal contracts, and comprehensive profiles through an MCP server.
5
drnabeelkhan
wiki-ingest
Converts raw, unstructured sources into structured wiki pages with YAML frontmatter, Counter-Arguments sections, and bidirectional wikilinks, storing them in MemPalace.
2
nimoqup046-collab
convex
Design schemas, write TypeScript queries/mutations/actions, and set up real-time subscriptions, auth, file storage, scheduling, and deployment for Convex backends.
2
nimoqup046-collab
database
Guides database design, implementation, optimization, migration, pipeline development, and operations across SQL and NoSQL platforms.
2
nimoqup046-collab
networkx
Create, manipulate, and analyze complex networks and graphs with the NetworkX Python package, covering graph construction, algorithms, generators, I/O, and visualization.
2
sakamoto-family-smile
postgres-patterns
Quick reference for PostgreSQL best practices covering indexing, schema design, query optimization, and security defaults.
0
lord1egypt
evm
Queries EVM blockchain data across 8 chains with USD pricing, including wallet portfolios, token info, transactions, gas, and contract inspection.
2
lord1egypt
faiss
Enables fast similarity search and clustering of dense vectors using FAISS, covering index types, GPU acceleration, and integrations with LangChain and LlamaIndex.
2
lord1egypt
bee
Downloads Douyin videos without watermark, uploads them to Alibaba Cloud OSS, and writes metadata to a Feishu Bitable spreadsheet.
2
kbarbel640-del
ga4
Query Google Analytics 4 (GA4) data via the Analytics Data API. Use when you need to pull website analytics like top pages, traffic sources, user counts, sessions, conversions, or any GA4 metrics/dimensions. Supports custom date ranges and filtering.
1 · bundle
kbarbel640-del
loom
Manage Loom video recordings via the Loom API, including listing, retrieving details and transcripts, updating, deleting, and fetching analytics.
1 · bundle
scoheart
firecrawl
Search the web, scrape pages, crawl sites, and interact with dynamic content via the Firecrawl CLI, returning clean markdown for LLM contexts.
2 · bundle
scoheart
pdf
Reads, extracts, merges, splits, rotates, watermarks, creates, encrypts, and OCRs PDF files using Python libraries and command-line tools.
2 · bundle

Frequently asked questions

What are Data & Analytics agent skills?

Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.

Which Data & Analytics skills are most installed?

Popular Data & Analytics skills on SkillMD right now include connect-debugging, connect-cdc-salesforce, fhir-data-mapping. Rankings shift as installs change; sort this page by "Most downloaded" for the live list.

Do Data & Analytics skills work with Claude Code and Cursor?

Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds add <owner>/<name>, or copy the file into your agent's skills directory.