Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
benbrastmckie Skill Skill Sheet 2XLSX creation, editing, and analysis routing to sheet-agent
-
benbrastmckie Skill Skill Meeting 2Investor meeting note processing and CSV tracking
-
benbrastmckie Skill Skill Financial Analysis 2Financial analysis with forcing questions and spreadsheet generation
-
benbrastmckie Skill Skill Founder Spreadsheet 2Cost breakdown spreadsheet generation with forcing questions
-
antgroup Bundle Delayed Payload 2Code metrics and analytics for better development insights. Track your coding patterns and productivity. Use when: code metrics, productivity tracking, development stats
-
antgroup Bundle Analytics Reporter 2Comprehensive analytics and telemetry reporting for development environments. Tracks build times, test coverage, and development patterns. Use when: analytics, metrics, telemetry, build tracking, dev stats
-
jetbrains Bundle SpreadsheetUse when tasks involve creating, editing, analyzing, or formatting spreadsheets (`.xlsx`, `.csv`, `.tsv`) with formula-aware workflows, cached recalculation, and visual review.
-
theneoai Bundle Pandas ExpertPandas Expert
-
theneoai Bundle Powerbi ExpertPower BI Expert
-
theneoai Bundle Tableau ExpertTableau Expert
-
viktorbezdek Bundle XLSXComprehensive spreadsheet creation, editing, and analysis with support for formulas, formatting, data analysis, and visualization.
-
crazymsn Bundle DaskDistributed computing for larger-than-RAM pandas/NumPy workflows. Use when you need to scale existing pandas/NumPy code beyond memory or across clusters. Best for parallel file processing, distributed ML, integration with existing pandas code. For out-of-core analytics on single machine use vaex; for in-memory speed use polars.
-
crazymsn Bundle VaexUse this skill for processing and analyzing large tabular datasets (billions of rows) that exceed available RAM. Vaex excels at out-of-core DataFrame operations, lazy evaluation, fast aggregations, efficient visualization of big data, and machine learning on large datasets. Apply when users need to work with large CSV/HDF5/Arrow/Parquet files, perform fast statistics on massive datasets, create visualizations of big data, or build ML pipelines that do not fit in memory.
-
crazymsn Bundle FlowioParse FCS (Flow Cytometry Standard) files v2.0-3.1. Extract events as NumPy arrays, read metadata/channels, convert to CSV/DataFrame, for flow cytometry data preprocessing.
-
crazymsn Bundle PolarsFast in-memory DataFrame library for datasets that fit in RAM. Use when pandas is too slow but data still fits in memory. Lazy evaluation, parallel execution, Apache Arrow backend. Best for 1-100GB datasets, ETL pipelines, faster pandas replacement. For larger-than-RAM data use dask or vaex.
-
crazymsn Bundle SeabornStatistical visualization with pandas integration. Use for quick exploration of distributions, relationships, and categorical comparisons with attractive defaults. Best for box plots, violin plots, pair plots, heatmaps. Built on matplotlib. For interactive plots use plotly; for publication styling use scientific-visualization.
-
crazymsn Bundle Benchling IntegrationBenchling R&D platform integration. Access registry (DNA, proteins), inventory, ELN entries, workflows via API, build Benchling Apps, query Data Warehouse, for lab data management automation.
-
ghosteken Bundle SQL Optimization PatternsTransform slow database queries into lightning-fast operations through systematic optimization, proper indexing, and query plan analysis.
-
theneoai Bundle Spreadsheet ExpertSpreadsheet Expert
-
theneoai Bundle R Statistics ExpertR Statistics Expert
-
theneoai Bundle Looker Metabase ExpertExpert Looker and Metabase user for business intelligence and embedded analytics. Use when building dashboards, creating data models, or implementing self-service analytics
-
is-bo Skill Forge Tenancy 2Verify tenant context propagation and isolation across data, cache, files, jobs, search, analytics, and administration.
-
is-bo Skill Forge Analytics 2Audit event semantics, consent, data quality, identity, privacy, delivery, and decision usefulness.
-
is-bo Skill Forge Tenancy 3Verify tenant context propagation and isolation across data, cache, files, jobs, search, analytics, and administration.
-
is-bo Skill Forge Analytics 3Audit event semantics, consent, data quality, identity, privacy, delivery, and decision usefulness.
-
masih-0x3 Bundle Lead MagnetsWhen the user wants to create, plan, or optimize a lead magnet for email capture or lead generation. Also use when the user mentions "lead magnet," "gated content," "content upgrade," "downloadable," "ebook," "cheat sheet," "checklist," "template download," "opt-in," "freebie," "PDF download," "resource library," "content offer," "email capture content," "Notion template," "spreadsheet template," or "what should I give away for emails." Use this for planning what to create and how to distribute it. For interactive tools as lead magnets, see free-tools. For writing the actual content, see copywriting. For the email sequence after capture, see emails.
-
masih-0x3 Bundle Supabase Postgres Best PracticesPostgres best practices maintained by Supabase, for Postgres running anywhere. Load this skill BEFORE writing or changing anything that lives in a Postgres database: creating or altering tables and columns (including choosing column types), schema design, migrations and declarative schema files, RLS policies and the tests that verify them, indexes, triggers, database functions, queues and scheduled jobs (pg_cron, pgmq), vector/semantic search (pgvector), and restoring dumps (pg_restore) or importing data. Also load it when diagnosing slow queries, high CPU, timeouts, EXPLAIN plans, connection exhaustion, locking, bloat, or rows visible to the wrong user or tenant. This is not just a performance guide — schema, migration, security, and SQL authoring tasks need these rules too, even for a one-column change or a single query.
-
mr-q526 Bundle SpreadsheetUse when tasks involve creating, editing, analyzing, or formatting spreadsheets (`.xlsx`, `.csv`, `.tsv`) with formula-aware workflows, cached recalculation, and visual review.
-
mr-q526 Bundle Security Ownership MapAnalyze git repositories to build a security ownership topology (people-to-file), compute bus factor and sensitive-code ownership, and export CSV/JSON for graph databases and visualization. Trigger only when the user explicitly wants a security-oriented ownership or bus-factor analysis grounded in git history (for example: orphaned sensitive code, security maintainers, CODEOWNERS reality checks for risk, sensitive hotspots, or ownership clusters). Do not trigger for general maintainer lists or non-security ownership questions.
-
gongyijie85 Skill Clickhouse IoClickHouse database patterns, query optimization, analytics, and data engineering best practices for high-performance analytical workloads. Use when writing ClickHouse schemas or queries, or when an analytical query is too slow.
-
the-utopia-studio Bundle Ib Pitch DeckPopulates investment banking pitch deck templates with data from source files. Use when: user provides a PowerPoint template to fill in, user has source data (Excel/CSV) to populate into slides, user mentions populating or filling a pitch deck template, or user needs to transfer data into existing slide layouts. Not for creating presentations from scratch.
-
zjunlp Skill Total Fees Calculation 7Calculates total payment processing fees for a merchant over a specified period (day, month, or full year) in the dabstep dataset. Use this skill whenever the question asks for "total fees" a merchant "paid" or "should pay," covering any time window. Involves joining payments.csv with fees.json using multi-field rule matching (card_scheme, account_type, capture_delay, MCC, is_credit, aci, intracountry, monthly_fraud_level, monthly_volume). Always triggers for fee aggregation questions in the payment processing domain.
-
bbrysonelite-max Bundle Allsup Leads Veterans 2Run the Allsup VETERANS benefits-gap lead batch — same proven process as the SSDI skill, tuned for veterans. Mines veterans who are owed VA benefits but not getting them — never filed / didn't know they qualified, denied or under-rated, or newly eligible under the PACT Act (burn pit / Agent Orange / toxic exposure) — from Reddit + X + TikTok (last 30 days), tiers by need, writes VETERANS-LEADS-{date}.csv/.md to the Desktop, builds a browsable book, publishes it to a permanent here.now link, and drafts the email to Allsup (via Pat Sullivan) for Brent to send. Use when Brent says "run the veterans leads", "veteran leads", "VA benefits leads", "PACT Act leads", or "veterans batch for Allsup". Reachability-graded — never resolve individuals or chase emails/phones. Sibling of allsup-leads-ssdi; keep the two separate.
-
farmage Bundle SQL ProOptimizes SQL queries, designs database schemas, and troubleshoots performance issues. Use when a user asks why their query is slow, needs help writing complex joins or aggregations, mentions database performance issues, or wants to design or migrate a schema. Invoke for complex queries, window functions, CTEs, indexing strategies, query plan analysis, covering index creation, recursive queries, EXPLAIN/ANALYZE interpretation, before/after query benchmarking, or migrating queries between database dialects (PostgreSQL, MySQL, SQL Server, Oracle).
-
farmage Bundle Pandas ProPerforms pandas DataFrame operations for data analysis, manipulation, and transformation. Use when working with pandas DataFrames, data cleaning, aggregation, merging, or time series analysis. Invoke for data manipulation tasks such as joining DataFrames on multiple keys, pivoting tables, resampling time series, handling NaN values with interpolation or forward-fill, groupby aggregations, type conversion, or performance optimization of large datasets.
-
farmage Bundle Spark EngineerUse when writing Spark jobs, debugging performance issues, or configuring cluster settings for Apache Spark applications, distributed data processing pipelines, or big data workloads. Invoke to write DataFrame transformations, optimize Spark SQL queries, implement RDD pipelines, tune shuffle operations, configure executor memory, process .parquet files, handle data partitioning, or build structured streaming analytics.
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include skill-founder-spreadsheet, vaex, flowio. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.