Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
aradotso Skill Enterprise User Management AI AnalyticsEnterprise user management system with AI-powered analytics for risk detection, burnout analysis, and predictive insights
-
aradotso Skill Coinmarketcap Diamonds Premium AnalyticsCoinMarketCap Diamonds premium build with unlocked pro analytics and trading features for cryptocurrency market analysis
-
nvidia-tao Bundle Tao Mine Aoi ImagesRuns the DEFT embed-then-mine workflow for VCN AOI iterations — embeds the gap-analysis target parquet, embeds a source pool, and mines nearest-neighbour source images for downstream augmentation. Use as the immediate next step after `tao-route-visual-changenet-samples` when expanding a real-image augmentation queue from the mining subset.
-
frumu-ai Skill Data Write QueryWrite optimized SQL for your dialect with best practices
-
aradotso Skill Commercevault Edd Ecommerce OrchestratorMaster the CommerceVault middleware for Easy Digital Downloads API integration, sales analytics, and digital commerce orchestration
-
aradotso Skill Harvard Artifacts Data Engineering StreamlitBuild end-to-end data engineering pipelines with Harvard Art Museums API, ETL workflows, SQL analytics, and Streamlit dashboards
-
aradotso Skill Enterprise User Management System AI AnalyticsFull-stack user management system with AI-powered analytics for risk detection, burnout analysis, and predictive insights
-
aradotso Skill Harvard Art Museums Data Engineering AnalyticsBuild ETL pipelines and analytics dashboards using the Harvard Art Museums API with Python, SQL, and Streamlit
-
frumu-ai Skill Bio Instrument DataConvert laboratory instrument output files (PDF, CSV, Excel, TXT) to Allotrope Simple Model (ASM) JSON format or flattened 2D CSV. Use this skill when scientists need to standardize instrument data for LIMS systems, data lakes, or downstream analysis. Supports auto-detection of instrument types. Outputs include full ASM JSON, flattened CSV for easy import, and exportable Python code for data engineers. Common triggers include converting instrument files, standardizing lab data, preparing data for upload to LIMS/ELN systems, or generating parser code for production pipelines.
-
dtsong Skill Generate PlotUse this skill when creating a statistical plot or chart from a data file. Triggers on "plot this data", "make a chart", "graph this CSV", or "visualize these results". Applies to CSV, JSON, or tabular data needing bar charts, scatter plots, line graphs, or similar visualizations. Do NOT use for methodology diagrams from text (use generate-diagram) or diagram scoring (use evaluate-diagram).
-
dtsong Skill Generate DiagramUse this skill when creating a methodology diagram from research text. Triggers on "make a diagram", "visualize this methodology", "diagram this process", or "generate a figure from this paper". Applies to methodology descriptions, process flows, and research paper sections. Do NOT use for scoring existing diagrams (use evaluate-diagram) or plotting data from CSV/JSON (use generate-plot).
-
oaustegard Bundle Converting FilesConvert a file from one format to another inside the container — documents, images, audio, video. Routes to the right engine (pandoc, LibreOffice, ImageMagick, ffmpeg) by format pair. Triggers on "convert X to Y", "turn this docx into a pdf", "make a gif from this mp4", "md to docx", "batch-convert these images", or any single-file or batch format change where the source and target extensions differ. NOT for editing content (use docx/pptx/xlsx/pdf skills), creating files from scratch, or reading a file you already have in context.
-
oaustegard Bundle Developing PreactSpecialized Preact development skill for standards-based web applications with native-first architecture and minimal dependency footprint. Use when building Preact projects, particularly those involving data visualization, interactive applications, single-page apps with HTM syntax, Web Components integration, CSV/JSON data parsing, WebGL shader visualizations, or zero-build solutions with vendored ESM imports.
-
agentuity-docs Skill Agentuity CLI Project Auth GenerateGenerate SQL schema for Agentuity Auth tables. Use for managing authentication credentials
-
oaustegard Bundle Reading Business CardsPreprocesses photographed sheets of many business cards — slicing each into overlapping high-resolution tiles and de-glaring them with container tooling (OpenCV/ImageMagick) — then reads every card via cheap parallel temperature-0 API calls (Haiku or Sonnet) using a distilled extraction prompt, and writes deduped contact fields to a CSV. Use when a user has photos or scans holding multiple business cards per image, mentions glare or unreadable cards, batch card transcription, contact extraction, or wants to read many cards without an expensive in-conversation pass. Triggers on 'business cards', 'card scan', 'extract contacts', 'read these cards', 'card glare', 'too many cards per photo'.
-
dtsong Skill Schema EvaluationEvaluate and design data warehouse schemas — star, snowflake, data vault, OBT — with grain definition, SCD strategies, and normalization trade-offs
-
dtsong Skill Analytics DesignTelemetry events, A/B test instrumentation, and success metrics design
-
dtsong Skill Impact EstimationRICE scoring framework for evidence-based feature prioritization
-
overtimepog Skill Backend QueriesWrite efficient, secure database queries using ORMs or raw SQL, preventing N+1 problems, SQL injection, and performance issues. Use this skill when writing database queries, implementing data access layers, creating repository patterns, or optimizing query performance in service files, query builders, or data access objects. Apply this skill when using parameterized queries, implementing eager loading to avoid N+1 queries, selecting only needed columns, adding WHERE/JOIN/ORDER BY clauses, or working with query optimization, indexes, and database performance tuning. This skill ensures queries use proper SQL injection prevention, implement transactions for data consistency, cache expensive queries appropriately, and follow best practices for query timeouts, connection pooling, and database resource management.
-
overtimepog Bundle XLSXSpreadsheet toolkit (.xlsx/.csv). Create/edit with formulas/formatting, analyze data, visualization, recalculate formulas, for spreadsheet processing and analysis.
-
w95 Skill Qbr BuilderCreate Quarterly Business Reviews with account health scores, usage analytics, ROI analysis, expansion opportunities, and risk mitigation
-
w95 Skill Fsi Comps AnalysisBuild institutional-grade comparable company analyses with operating metrics, valuation multiples, and statistical benchmarking in Excel/spreadsheet format. **Perfect for:** - Public company valuation (M&A, investment analysis) - Benchmarking performance vs. industry peers - Pricing IPOs or funding rounds - Identifying valuation outliers (over/under-valued) - Supporting investment committee presentations - Creating sector overview reports **Not ideal for:** - Private companies without comparable public peers - Highly diversified conglomerates - Distressed/bankrupt companies - Pre-revenue startups - Companies with unique business models
-
dtsong Bundle Python Data EngineeringUse this skill when writing Python code for data pipelines or transformations. Covers Polars, Pandas, PySpark DataFrames, dbt Python models, API extraction scripts, and data validation with Pydantic or Pandera. Common phrases: "Polars vs Pandas", "PySpark DataFrame", "validate this data", "Python extraction script". Do NOT use for SQL-based dbt models (use dbt-transforms) or integration architecture (use data-integration).
-
arabelatso Bundle Taint Instrumentation AssistantInstruments code to track the flow of untrusted or sensitive data at runtime, enabling detection of injection vulnerabilities, data leaks, and privilege violations. Use when users need to: (1) Track untrusted input propagation through code, (2) Detect SQL injection, XSS, or command injection vulnerabilities, (3) Identify sensitive data leaks, (4) Monitor privilege escalation paths, (5) Perform dynamic taint analysis for security testing. Supports Python, Java, JavaScript, and C/C++ with configurable taint sources and sinks.
-
w95 Skill User Research SynthesizerSynthesize user research findings from interviews, surveys, and analytics. Create insight reports, customer journey maps, and actionable recommendations based on research data and qualitative findings.
-
clawdsolana Bundle DatabaseCreate and manage Replit's built-in PostgreSQL databases, check status, execute SQL queries with safety checks, and run read-only queries against the production database. Use when the user wants to check prod data, debug database issues in production, or asks to "check the prod db", "query production", "look at live data", or "see what's in the database on the deployed app". Also use when the user asks how to apply development schema changes to the production database, e.g. "push dev to prod", "migrate the production database", "sync the schema", "production is missing a column", or reports a deployed app failing with "column does not exist" / "relation does not exist".
-
clawdsolana Skill Pump MCP ServerModel Context Protocol server exposing 53 tools, 3 resource types, and 3 prompts for AI agent consumption — quoting, building transactions, fee management, analytics, AMM operations, social fees, wallet operations over stdio transport.
-
masharratt Bundle Codesearch Code SearchMANDATORY: Query CodeSearch BEFORE using grep, glob, find, or search. 400x faster semantic/structural code search via SQL on indexed codebase. Use to find functions, classes, patterns, callers, implementations. Agents MUST query CodeSearch first; grep only allowed after CodeSearch returns zero results.
-
clawdsolana Bundle Query Integration DataQuery and modify data in any connected integration (Linear, GitHub, HubSpot, Slack, Google services, etc.) or connected data warehouse (Databricks, Snowflake, BigQuery). Use listConnections() in the code_execution sandbox to get credentials, then call APIs directly. Supports read operations (queries, counts, exports) and write operations (create, update, delete).
-
ricable Bundle Ruvector RvliteStandalone vector database with SQL, SPARQL, and Cypher query support powered by RuVector WASM. Use when the user needs a lightweight embedded vector database, multi-language query support (SQL/SPARQL/Cypher), standalone vector search without external dependencies, or a portable vector store for applications.
-
clawdsolana Bundle Excel GeneratorCreate Excel spreadsheets with formulas, charts, pivot summaries, and financial models.
-
masharratt Bundle Cfn Parameterized QueriesParameterized SQL execution, blocks injection. Use when executing DB queries, inserting/updating records, or any SQL op needing security hardening.
-
rjmurillo Bundle Software Engineering LibraryRoute software engineering design and discovered code-risk tasks to on-demand book references. Use for `architecture review`, `layer boundary change`, `dependency boundary`, `module interface shape`, `domain modeling`, `bounded context`, `refactoring`, `code smell`, `legacy code`, `low test coverage`, `old file`, `characterization test`, `external API calls`, `queues`, `retries`, `transactions`, `event ordering`, `data layer`, `storage design`, `consistency`, `schema evolution`, `timeout`, `circuit breaker`, `bulkhead`, and production resilience in .py, .cs, .ts, .tsx, .js, .ps1, .sql, and service design docs. Do NOT use for reinventing-the-wheel or build-vs-buy, use programming-advisor. Do NOT use for single-file maintainability scoring, use code-qualities-assessment. Do NOT use for CVA design, use cva-analysis.
-
ricable Bundle Ruvector Edge FullComplete WASM edge toolkit: vector search, graph DB, neural networks, DAG workflows, SQL/SPARQL/Cypher, ONNX inference. Use when the user needs an all-in-one edge AI runtime with vector search, graph queries, neural inference, task scheduling, or multi-query-language support in browser or edge environments.
-
fierzone Bundle Golang SecuritySecurity standards for Go backend services (Input Validation, Crypto, SQL Injection Prevention).
-
ricable Bundle Ruvector Postgres CLIPostgreSQL AI vector database CLI with pgvector-compatible extension, 53+ SQL functions, HNSW/GNN/attention ops. Use when the user needs to manage PostgreSQL vector operations, install the RuVector extension, run vector/sparse/hyperbolic/graph/attention/GNN queries, benchmark PostgreSQL vector performance, or manage a PostgreSQL-backed AI database.
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include agentuity-cli-project-auth-generate, enterprise-user-management-ai-analytics, coinmarketcap-diamonds-premium-analytics. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.