Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
liaosw97 Bundle Data WarehouseUse when working with data warehouses — ETL, dimensional modeling, analytics, data pipelines
-
liaosw97 Skill Mysql SQL RulesMySQL SQL语句规约 — 查询、更新、分页优化。Use when writing SQL queries, optimizing performance, handling pagination.
-
liaosw97 Bundle Python Data ProcessingUse when working with Python data processing — ETL, data pipelines, type-safe data handling
-
liaosw97 Bundle Java Mysql DatabaseUse when working with MySQL in Java — database schema, SQL, ORM mapping
-
zuojuzhang Bundle Business Analyst面向决策者的商业分析师 skill。当用户提到分析数据、分析、数据分析、分析数据、商业分析、策略分析、ROI分析、投放效果、效率评估、优化建议、看数据、分析表格等关键词时触发。支持 Excel/CSV/文本数据输入,输出包含 Actionable Insight、Key Metrics、可视化图表的专业分析报告。即使用户只是说"帮我看看这个数据"或"分析一下这份数据",也应当使用此技能。
-
liaosw97 Bundle Data Science Pandas Best PracticesUse when working with Pandas — DataFrame operations, data manipulation, best practices
-
macrossgithub-coder Bundle Hive Doris SQL Transformation将 Hive HQL 文件转换为 Apache Doris SQL 文件。支持单文件和目录批量处理。触发场景:用户提供 .hql 文件路径或目录路径,要求转换/迁移 Hive SQL 到 Doris;或用户说"把这个 HQL 转成 Doris SQL"、"迁移 Hive 表到 Doris"、"批量转换 HQL 文件"等。
-
pv56bkd4y8-design Skill AI Literacy Build Eval DatasetAssembles Coze Loop (扣子罗盘) eval CSV/JSONL for 结果评分v2 from a manifest of answer types and generated deliverable files. Use when building or refreshing the result-scoring eval set after generating candidate outputs with ai-literacy-generate-deliverable.
-
yrzhe Bundle Akshare A Shares针对中国A股的行情与数据查询(基于本地 akshare 源代码与API映射)。当用户请求获取A股实时行情、历史K线(日/周/月)、分钟级分时(1/5/15/30/60)、复权数据(前复权/后复权)、股东户数(含详情)、以及巨潮资讯公告披露结果时使用本Skill。默认优先使用东方财富(EM)接口;如用户指定新浪(Sina)来源或需要特定接口,则按映射调用对应函数。可使用内置脚本输出CSV/JSON,或内联Python调用。
-
moii-dev Bundle Student Homework BuilderUse this skill whenever the user asks Codex to complete, improve, structure, or prepare a school/college programming assignment, lab work, homework project, database task, web assignment, Android task, ASP.NET task, Python task, C# task, SQL task, or educational project. The skill ensures the result follows the teacher's requirements, stays understandable for a student, avoids unnecessary overengineering, and is easy to explain during defense.
-
ila Skill Storage FormatsReference for data lake storage formats and table formats — Parquet, Delta Lake, Apache Iceberg, Apache Hudi, DuckLake, Lance, F3 (Faster Frozen Format), ORC, and their interaction with DuckDB. Auto-loaded when discussing storage formats, data lakes, table formats, Parquet, Iceberg, Delta Lake, DuckLake, or change data capture.
-
xiangdong-415 Bundle Office Spreadsheet WorkflowWorkflow guidance for Codex agents creating, editing, analyzing, visualizing, validating, or formatting Microsoft Excel, XLSX, CSV, TSV, Office spreadsheet, and Google Sheets-targeted files through available spreadsheet tools or MCP/plugin integrations. Use for Excel models, tables, dashboards, formulas, charts, data cleaning, research data workbooks, financial models, trackers, and Chinese requests about Excel, xlsx, biaoge, shuju fenxi, gongshi, or tubiao.
-
ila Skill Write DocsStyle guide for writing OpenIVM documentation. Auto-loaded when writing, editing, or reviewing docs, README, user guides, or SQL reference pages. Blends Snowflake's structured clarity with DuckDB's concise, example-first approach.
-
tcvdog Skill Document GeneratorExpert document creation specialist who generates professional PDF, PPTX, DOCX, and XLSX files using code-based approaches with proper formatting, charts, and data visualization.
-
tcvdog Skill Analytics ReporterExpert data analyst transforming raw data into actionable business insights. Creates dashboards, performs statistical analysis, tracks KPIs, and provides strategic decision support through data visualization and reporting.
-
tcvdog Skill Retail Customer ReturnsComprehensive retail customer returns specialist for processing returns, exchanges, and refunds across in-store, online, and omnichannel retail — handling policy enforcement, fraud prevention, customer retention, vendor returns, and returns analytics to maximize recovery while preserving customer loyalty
-
tcvdog Skill Sales Data Extraction AgentAI agent specialized in monitoring Excel files and extracting key sales metrics (MTD, YTD, Year End) for internal live reporting
-
tuananhcr Bundle Universal File ConverterChuyển đổi file thông minh giữa tất cả các định dạng phổ biến trực tiếp trong chat. AI tự chuyển đổi nội dung dạng text (PDF->MD, CSV->JSON, etc.) mà không cần sandbox. Với các định dạng nhị phân/cần render (->DOCX, ->PDF, ->XLSX), AI sẽ hướng dẫn giải pháp thay thế thông minh nhất mà không cố chạy sandbox vô ích.
-
luokai0 Bundle Vaex 2Use this skill for processing and analyzing large tabular datasets (billions of rows) that exceed available RAM. Vaex excels at out-of-core DataFrame operations, lazy evaluation, fast aggregations, efficient visualization of big data, and machine learning on large datasets. Apply when users need to work with large CSV/HDF5/Arrow/Parquet files, perform fast statistics on massive datasets, create visualizations of big data, or build ML pipelines that do not fit in memory.
10 -
luokai0 Bundle Seaborn 2Statistical visualization with pandas integration. Use for quick exploration of distributions, relationships, and categorical comparisons with attractive defaults. Best for box plots, violin plots, pair plots, heatmaps. Built on matplotlib. For interactive plots use plotly; for publication styling use scientific-visualization.
10 -
luokai0 Skill Claimable Postgres 2Provision instant temporary Postgres databases via Claimable Postgres by Neon (neon.new) with no login, signup, or credit card. Supports REST API, CLI, and SDK. Use when users ask for a quick Postgres environment, a throwaway DATABASE_URL for prototyping/tests, or "just give me a DB now". Triggers include: "quick postgres", "temporary postgres", "no signup database", "no credit card database", "instant DATABASE_URL", "npx neon-new", "neon.new", "neon.new API", "claimable postgres API".
10 -
luokai0 Bundle Timesfm Forecasting 2Zero-shot time series forecasting with Google's TimesFM foundation model. Use for any univariate time series (sales, sensors, energy, vitals, weather) without training a custom model. Supports CSV/DataFrame/array inputs with point forecasts and prediction intervals. Includes a preflight system checker script to verify RAM/GPU before first use.
10 -
ecnu-icalk Skill Excel 9根据用户指定的步骤编写Excel VBA宏,执行筛选首行、删除空白行、清洗指定列数据并求和的操作。
559 -
ecnu-icalk Skill CSV 5当用户请求地铁站或地点的经纬度信息时,按照指定的字段顺序(站口名、经度、纬度)和逗号分隔符输出数据,并去除行号前缀。
559 -
luokai0 Bundle Clinical Decision Support 2Generate professional clinical decision support (CDS) documents for pharmaceutical and clinical research settings, including patient cohort analyses (biomarker-stratified with outcomes) and treatment recommendation reports (evidence-based guidelines with decision algorithms). Supports GRADE evidence grading, statistical analysis (hazard ratios, survival curves, waterfall plots), biomarker integration, and regulatory compliance. Outputs publication-ready LaTeX/PDF format optimized for drug development, clinical research, and evidence synthesis.
10 -
digitalexplorers Skill PHP Security AuditAudits PHP codebases for critical security vulnerabilities including SQL injection, XSS, command injection, file inclusion, insecure deserialization, weak cryptography, CSRF, and misconfigured PHP settings
-
ivanshamaev Skill Trino Dbt Query PerformanceOptimizing dbt-generated SQL performance on Trino — ephemeral vs view vs table materialization trade-offs, CTE explosion patterns, partition-aware incremental filters (bounded watermarks), MERGE vs delete+insert strategy selection, avoiding full-refresh anti-patterns, query hints in dbt SQL (BROADCAST hint), dbt thread tuning, session_properties for dbt runs, ANALYZE post-hooks, incremental model design patterns for Iceberg, avoiding small file proliferation from frequent incremental runs
-
ivanshamaev Skill Trino Observability PlatformTrino observability and monitoring platform — JMX Prometheus exporter configuration (running queries/failed queries/OOM kills/execution latency P50/P90/P99/memory pool metrics), Grafana dashboard panels, OpenTelemetry trace propagation, query-level event listener for structured logging, Prometheus alert rules (worker loss/queue depth/OOM/failure rate/p99 latency), log aggregation patterns, query history analysis via REST API, slow query detection SQL
-
ivanshamaev Skill Dbt Starrocks Performancedbt + StarRocks performance tuning — model DAG optimization (fan-in/fan-out), incremental merge window sizing, partition-aware incremental filters, avoiding full table rebuilds, query plan hints in dbt SQL (LEADING/JOIN hints), materialized view as dbt model target, concurrent model execution with threads, pre/post hooks for ANALYZE TABLE, dbt run selectors to minimize rebuilt models
-
ivanshamaev Skill Starrocks Files IngestionStarRocks file ingestion — FILES() table function (SELECT/INSERT from S3/HDFS Parquet/ORC/CSV without CREATE TABLE), Iceberg external catalog (HMS/Glue/REST), CREATE EXTERNAL CATALOG, cross-catalog INSERT INTO SELECT, schema auto-detection, partition filter pushdown on external tables, SHOW CREATE CATALOG, external table DDL patterns
-
lypersonalgit Skill Flomo Tag Analyzer分析和优化 Flomo 或任何基于标签的笔记工具的标签分类体系。 执行标签层级分析、命名规范检查、分类逻辑评估,生成优化后的标签结构和迁移方案。 使用场景: - 分析或优化 Flomo/笔记应用的标签体系 - 整理混乱、重复或难以检索的标签 - 搭建符合 PARA 体系或个人习惯的标签结构 - 合并、重命名、调整现有标签层级 - 将优化结果导出为可导入的 .md/.csv 文件 - 标签使用报告,识别冷门/孤立标签 - 优化标签、标签分析、标签整理、标签重构、Flomo标签、tag analysis
-
ivanshamaev Skill Trino File Layout OptimizationTrino data file layout optimization for Iceberg — Parquet vs ORC file format selection, target file size tuning (iceberg.target-max-file-size), row group size, Parquet/ORC column encoding choices, Bloom filter indexes, sorted_by for min/max skipping, small file detection via $files metadata table, OPTIMIZE compaction strategies, partition design impact on file count, Z-order equivalent via sorted_by, split sizing and parallelism (iceberg.minimum-assigned-split-weight), write parallelism tuning
-
ivanshamaev Skill Starrocks AI Query AutotunerStarRocks AI query autotuner — autonomous SQL optimization agent workflow (EXPLAIN COSTS → analyze → recommend), materialized view recommendation from slow query log, index recommendation (bitmap/bloom filter), partition pruning diagnosis, join order hints generation, statistics staleness detection and auto-ANALYZE trigger, slow query pattern classification, query rewrite suggestions
-
ivanshamaev Skill Starrocks Realtime AnalyticsStarRocks real-time analytics — Kafka → Routine Load → Primary Key table for sub-second freshness, low-latency BI query patterns, real-time dashboard design (materialized view for pre-aggregation), metric store patterns, window-based freshness for streaming dashboards, colocate join for real-time multi-table queries, resource group isolation for OLAP vs ingestion
-
ivanshamaev Skill Starrocks Routine Load KafkaStarRocks Routine Load for Kafka — CREATE ROUTINE LOAD DDL (all PROPERTIES/KAFKA clause parameters), JSON/CSV/Avro format config, desired_concurrent_number tuning, SHOW ROUTINE LOAD columns, PAUSE/RESUME/ALTER/STOP, consumer lag monitoring, error log analysis, exactly-once semantics, Schema Registry for Avro, Kafka SASL/SSL config, idempotent upsert to Primary Key table
-
ivanshamaev Skill DuckdbDuckDB in-process OLAP analytics — reading Parquet/CSV/JSON/Iceberg/Delta from S3/local, SQL (window functions, PIVOT, ASOF join, QUALIFY), Python API (duckdb.connect, fetchdf, register, UDFs), extensions (httpfs/iceberg/delta/postgres), performance tuning, COPY TO, persistent DB
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include data-warehouse, mysql-sql-rules, python-data-processing. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.