Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
rdoser13 Skill Linkedin Profile ScraperScrape LinkedIn profile post data using the Apify LinkedIn Profile Post Scraper and export it as a structured CSV. Use this skill when the user wants to pull LinkedIn post metrics, analyze LinkedIn content performance, scrape a LinkedIn profile's posts, or export LinkedIn engagement data. Also trigger when the user mentions Apify and LinkedIn together, or asks about post-level LinkedIn analytics.
-
vaquarkhan Skill Skill 07 Token AnalyticsAnalyzes on-chain token data including whale tracking, DEX volume, wallet PnL, and token unlock schedules. Use when researching token metrics, portfolio performance, or market activity.
-
vermapragya Bundle Cohort AnalysisBuilds cohort retention tables and retention curves with a single consistent denominator policy. Use when the user mentions cohort, retention, retention curve, "by signup month", "by acquisition channel", N-day retention, churn over time, or lifecycle analysis.
-
vermapragya Bundle Modular SQL CtesRefactors SQL into staging, intermediate, and fact CTE layers with explicit grain and naming conventions. Use when the user asks to refactor a SQL query, clean up a model, build a dbt model, modularize a query, or mentions CTE structure, query readability, or "this SQL is hard to follow."
-
vermapragya Bundle SQL Query ReviewStatic review of a SQL query — finds anti-patterns, structural problems, and performance issues from the query text alone, then proposes an optimized rewrite with a verification plan. Use when the user says "review this query", "check my SQL", "is this query okay", "clean up this query", or "optimize this" without runtime profile data.
-
vermapragya Bundle SQL Correctness ReviewAudits a SQL query's logic for wrong-results bugs — duplicate rows, join fanout, wrong join types, NULL handling traps, and CASE branch issues — with evidence queries that prove or clear each suspicion. Use when the user says "numbers look wrong", "double counting", "rows are duplicated", "this join is exploding", "check my query logic", or when totals don't reconcile.
-
vermapragya Bundle Warehouse Query OptimizationDiagnoses and fixes slow Snowflake queries — clustering, partition pruning, joins, spilling, warehouse sizing, and query plan reading. Use when the user mentions slow query, query optimization, Snowflake performance, query plan, clustering, micro-partitions, spilling, warehouse cost, or "this query is taking forever."
-
brianbolze Bundle Query CompaniesUse for named-company comparison or lookup prompts that can be answered from the local web-research company store, especially pricing, offerings, ownership, catalog breadth, cohort cuts, price-visibility, and broad single-company briefs: "tell me about X", "compare X and Y", "what do these competitors charge / offer", "which brands sell <thing>", "who owns X", "is X in the web-research store", "what has the store captured on X". Answers are captured-state snapshots from cited primary-source dossiers instead of WebSearch. Do not browse, WebSearch, curl, or open live company pages just to verify; resolve with store.py find, answer from store files, and cite local paths plus governing capture clocks. Read-only: never scrapes, never spends. If current/latest or missing/stale capture is required, suggest /research-company refresh; for incomplete rosters suggest /deepen-offerings; for external signals use tools/. Not for general web search, news/funding events, financials, judgments, or non-company topics.
-
clawdotnet Skill Data AnalystConnects to databases, runs SQL queries, and analyzes datasets using code to provide actionable business insights.
-
iwritec0de Bundle Wordpress SecurityThis skill should be used when the user asks to "secure a WordPress plugin", "escape output", "sanitize input", "verify nonces", "check user capabilities", "write safe database queries", or mentions "WordPress security", "esc_html", "sanitize_text_field", "wp_nonce", "wpdb prepare", "XSS", "CSRF", "SQL injection", "capability check", "wp_kses". Provides WordPress security best practices including output escaping, input sanitization, nonce verification, capability checks, and secure database queries.
-
juneqqq Skill Add Card TypeStandardized flow for adding a new message card type (13th payload kind). Use when adding rich UI cards beyond the existing 12 (text/diff/web/tasks/swatches/copy/metrics/sql/schema/logs/api/typing/ask-form).
-
juneqqq Skill Excel AnalystBuild useful spreadsheets, financial models, formulas, pivots, and editable xlsx outputs.
-
materializeinc Skill Mz BenchmarkAdd/modify/debug Materialize perf benchmark scenarios. Three frameworks: Feature Benchmark (single-op micro), Scalability Test (SQL throughput under concurrency), Parallel Benchmark (sustained latency via scenarios.py). Trigger: "benchmark", "feature benchmark", "scalability test", "parallel benchmark", "performance regression", "micro-benchmark", "TPS", "latency test", or edits in feature_benchmark/scenarios/, scalability/workload/workloads/, parallel_benchmark/scenarios.py. Note: measurement, not panic-stress (see mz-parallel-workload).
-
materializeinc Skill Mz Query PerfAnalyze and optimize a Console/catalog SQL query in user space — diagnose its plan on real relations, then measure candidate rewrites with a faithful synthetic-fleet sweep. Use when asked to find performance improvements for a query that reads mz_catalog / mz_internal relations.
-
materializeinc Bundle Mz Query TracingDebug SQL execution time via distributed tracing (OpenTelemetry / Tempo). Trigger: "why is this query slow", "where is the time going", "this SELECT takes forever", or latency breakdown for SQL statement. Also tracing queries, span analysis, Tempo traces, trace IDs, opentelemetry_filter.
-
materializeinc Skill Mz Parallel WorkloadExtend parallel-workload stress framework: random SQL concurrently to catch panics + unexpected errors (not perf — see mz-benchmark). Trigger: "parallel workload", "parallel-workload", "action.py" re parallel workload, or testing panics/unexpected errors under concurrency. Also "add this to parallel workload" or bug that panics under concurrent DDL/DML.
-
databricks-solutions Bundle Databricks GenieCreate and query Databricks Genie Spaces for natural language SQL exploration. Use when building Genie Spaces or asking questions via the Genie Conversation API.
-
databricks-solutions Bundle Databricks App PythonBuilds Python-based Databricks applications using Dash, Streamlit, Gradio, Flask, FastAPI, or Reflex. Handles OAuth authorization (app and user auth), app resources, SQL warehouse and Lakebase connectivity, model serving integration, and deployment. Use when building Python web apps, dashboards, ML demos, or REST APIs for Databricks, or when the user mentions Streamlit, Dash, Gradio, Flask, FastAPI, Reflex, or Databricks app.
-
databricks-solutions Bundle Databricks Agent BricksCreate and manage Databricks Agent Bricks: Knowledge Assistants (KA) for document Q&A, Genie Spaces for SQL exploration, and Supervisor Agents (MAS) for multi-agent orchestration. Use when building conversational AI applications on Databricks.
-
receiptor-ai Bundle Receipt ProcessingExtract structured data from receipts and invoices for bookkeeping, expense tracking, and tax preparation. Supports email, photos, scanned images, PDFs, and accounting exports. Outputs vendor, date, amount, tax, line items as table, CSV, or JSON. Trigger on "process receipts", "extract receipts from email", "scan invoices", "capture expenses".
-
svgreg Bundle PDF Table ExtractorExtracts tables from PDF files and outputs them as CSV. Use when the user asks to pull tabular data out of a PDF.
-
songhonglei Bundle Recover Codex Project ChatsDiagnose and safely repair Codex Desktop projects that exist but show “No chats/没有聊天”, while conversations may still appear under Recent. Use when users report missing project histories, lost thread-to-project mappings, model-provider changes, archived or invisible threads, moved cwd paths, incomplete ~/.codex restores, state_5.sqlite integrity or index problems, or need comparison against codex_threads_snapshot.csv or a ~/.codex backup.
-
grandcamel Bundle Jira Search JqlFind issues by criteria (status, assignee, priority, etc.) using JQL. Create filters, export results to CSV/JSON, bulk update. Ideal for reporting and automation.
-
graphistry Bundle PygraphistryTOC router for PyGraphistry Python SDK tasks. Use when asked to "plot a graph", "visualize my edges", "load a dataframe into graphistry", "import graphistry", "run UMAP on my graph", "query my graph with GFQL or Cypher", "pick an execution engine", "make my GFQL query faster", "connect graphistry to Neo4j/Splunk/Kusto", or any Python SDK graph workflow. Also triggers on "graphistry.register", "g.plot()", ".gfql()", "chain-list", "hypergraph", "engine='polars'", "polars-gpu", "cudf", or "pygraphistry". Dispatches to pygraphistry-core (auth/ETL/plot), pygraphistry-gfql (queries, engines, indexes), pygraphistry-visualization (styling/sharing), pygraphistry-ai (UMAP/DBSCAN/embeddings), or pygraphistry-connectors (external DBs). Proactively suggest when the user shares an edges/nodes DataFrame and asks about graph analysis or visualization.
-
graphistry Skill Pygraphistry CoreCore PyGraphistry workflow: auth, DataFrame-to-graph shaping, and first interactive plot. Use when asked to "register graphistry", "get started with pygraphistry", "plot my edges dataframe", "graphistry.register()", "bind src and dst columns", "make a hypergraph", "materialize nodes", or any first-graph / ETL-to-plot task. Also triggers on "first graphistry graph", "graphistry install", "api=3", or questions about graphistry auth credentials. Proactively suggest when the user is setting up graphistry for the first time or can't get a basic plot working from a DataFrame.
-
graphistry Skill Pygraphistry ConnectorsPyGraphistry connector workflows for external data sources and graph databases. Use when asked to "connect graphistry to Neo4j", "load from Splunk into graphistry", "query Kusto/ADX and visualize", "Databricks graph", "TigerGraph with pygraphistry", "ingest SQL into a graph", or any "graphistry + [external platform]" request. Also triggers on Neptune, Postgres, BigQuery, Memgraph, or connector/plugin keywords. Proactively suggest when the user has data in an external system and wants graph visualization without first loading it into a DataFrame.
-
lensetek Skill Desktop Rpa Computer UseSpesialis Otomatisasi Desktop OS & Computer Use untuk mengontrol GUI aplikasi pajak & akuntansi non-web (e-SPT Desktop, e-Faktur Client-Desktop, Accurate Desktop, MYOB, Zahir, Excel Macros) dengan strategi Primary & Auto-Fallback Recovery.
-
lubusin Bundle ReportsCreate reports in Frappe including Report Builder, Query Reports (SQL), and Script Reports (Python + JS). Use when building data analysis views, dashboards, or custom reporting features.
-
m3dcodie Skill Eng Flow AnalyticsProduction Stage 10 — on-demand rollup report of eng-flow/analytics.jsonl (time/token log) and eng-flow/findings.jsonl (bug counts), both incrementally written by other eng-flow skills. Reports per-story and per-stage totals, tokens by category, cycle time (elapsed vs. active), stage-transition gaps, review/QA rework cycles, and bug rate. Read-only; doesn't instrument itself.
-
mathruffian-dot Skill Opencode File Toolkit安裝 agent 的內部工具包三合一——A 文件處理(Word/Excel/PPT/PDF/圖片/QR/轉 Markdown 的 Python 標配 10 套件)、B 影音工具(yt-dlp/FFmpeg/deno)、C 語音(Edge-TTS 讓 agent 開口說話)。說「裝內部工具包」「裝教學檔案處理工具包」「裝 Python 檔案工具」「裝 yt-dlp」「裝影音下載工具」「讓 agent 會說話」「裝 Edge-TTS」「裝語音」時載入。
-
antongulin Skill API TestingAPI testing patterns for Playwright + TypeScript — resource class pattern (HTTP wrappers extending BasePage, domain folders mirroring REST namespaces, typed payload builders), the ApiListener pattern for capturing real responses by stateKey without mocking, optional SQL/stored-procedure bridge for test-data setup, and TOTP-based MFA enrollment. Use when seeding test data via API, asserting on API responses without mocking, building HTTP wrapper classes, capturing network responses during UI tests, or setting up test users with MFA.
-
arthurgailes Bundle R DuckplyrUse when code loads or uses duckplyr (library(duckplyr), duckplyr::), processing large datasets with dplyr syntax, working with Parquet files in R, or needing lazy evaluation for bigger-than-memory data
-
zjp1997720 Bundle Wxmp Article Harvester抓取、筛选并导出微信公众号公开文章的独立 Skill。用户提到微信公众号、公众号文章、mp.weixin、wcx、搜索公众号、导出最近 N 天或某年度文章、批量抓取、断点续抓、正文补全、教程文章筛选、微信文章链接保存时使用。支持 Markdown/JSON/CSV,默认用 wcx 获取索引、Playwright 提取正文;付费 Metaso 兜底必须显式授权。
-
cangyeone Bundle Tabular IoTabular Data Reading (CSV / TXT)
-
pandelisz Skill Codex Primary RuntimeContainer for Codex primary runtime skills including PowerPoint slides and Excel spreadsheets
-
pandelisz Bundle ExcelUse this skill when a user requests to create, modify, analyze, visualize, or work with spreadsheet files (`.xlsx`, `.xls`, `.csv`, `.tsv`) with formulas, formatting, charts, tables, and recalculation.
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include skill-07-token-analytics, query-companies, recover-codex-project-chats. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.