Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
damionrashford Bundle Media ScenedetectReliable scene-change detection with PySceneDetect (scenedetect CLI): content-aware detection (ContentDetector), threshold detection (ThresholdDetector), adaptive detection (AdaptiveDetector), scene list to CSV/JSON, split videos at scene cuts, save scene thumbnails, HSV/edge-based metrics. Use when the user asks to split a video at scene changes, build a chapter list from scene cuts, generate thumbnails per scene, detect shot boundaries reliably, or do content-aware auto-chapter markers (better than ffmpeg scdet).
-
stuffbucket Bundle Tauri PluginsUse when adding an official Tauri v2 plugin — picking the right plugin (fs/dialog/shell/http/store/notification/clipboard/global-shortcut/logging/os/opener/process/single-instance/autostart/deep-link/sql/websocket/upload/stronghold/cli), installing it (Cargo + npm), registering it in Rust, and granting the required capability permissions.
-
wenerme Bundle Doris DocsUse when working with Apache Doris: table design, data models (Duplicate/Unique/Aggregate), partitioning, bucketing, SQL syntax, data import (Stream Load, Broker Load, INSERT INTO), data export, lakehouse (Hive/Iceberg/Hudi/Paimon catalogs), materialized views, query acceleration, inverted index, compute-storage decoupled mode, administration, or Doris ecosystem tools.
-
wenerme Bundle Duckdb SQLUse when writing, debugging, optimizing, or checking DuckDB SQL syntax, statements, functions, data types, dialect, or query semantics.
-
wenerme Bundle Duckdb DataUse when importing, exporting, reading, writing, or bulk-loading CSV, JSON, Parquet, Iceberg, or other data in DuckDB.
-
wenerme Bundle Mikro Orm V6 To V7Use when upgrading @mikro-orm packages from v6 to v7, fixing v7 runtime/type errors (decorator SyntaxError, persistAndFlush removed, nativeInsert not found), adapting knex to kysely or better-sqlite to new SQLite drivers, running MikroORM in Edge/Bun/node:sqlite environments, or choosing between defineEntity vs decorator entity definitions. Triggers on "mikro-orm v7", "persistAndFlush", "@mikro-orm/decorators", "@mikro-orm/sql", "defineEntity", "bun:sqlite mikro-orm".
-
yunseo-kim Bundle Data AnalystSQL, pandas, and statistical analysis expertise for data exploration, cleaning, transformation, and insight generation
-
vignesh2027 Skill Crypto SageActivates CryptoSage for crypto, DeFi, and Web3 intelligence. Use when you need on-chain analytics (MVRV, SOPR, NVT, exchange flows), DeFi TVL trend analysis, tokenomics review (vesting schedules, inflation rate, unlock impact), narrative momentum tracking, or rug pull / audit risk assessment.
-
vignesh2027 Skill SQL AnalyzerActivates SQLAnalyzer for advanced SQL optimization, query analysis, and database performance tuning. Use when you need to optimize a slow query using EXPLAIN plans, rewrite subqueries as CTEs or window functions, design complex analytical queries, identify missing indexes, eliminate N+1 patterns, or write advanced SQL using window functions, recursive CTEs, or pivot logic.
-
vignesh2027 Skill Cfo IntelligenceActivates the CFO-Intelligence agent for financial analysis, reporting, and forecasting. Use this skill when you need to parse P&L, balance sheet, or cash flow statements; perform budget vs actual variance analysis with root cause; build 3-statement financial models with base/bull/bear scenarios; benchmark metrics vs industry comps; or generate board-ready executive summaries. Works with CSV, PDF, text, or pasted numbers.
-
vignesh2027 Skill Document ProcessorActivates DocumentProcessor for intelligent processing of Word, PDF, PowerPoint, and Excel files. Use when you need to extract, summarize, redline, or transform documents in any Office format — compare versions, extract tables, generate document outlines, convert formats, or produce structured data from unstructured documents.
-
thelobbi Bundle Office ScriptsExpert knowledge of Excel Office Scripts — Microsoft's TypeScript automation platform for Excel on the web, including the full ExcelScript API surface, TypeScript 4.0.3 restrictions, Graph API workbook integration, Power Automate connector limits, performance optimization, and common automation patterns.
-
thelobbi Bundle Pandas CleaningExpert knowledge of data cleaning with pandas — reading messy data from any source, transforming it into clean DataFrames, and outputting polished .xlsx files with openpyxl formatting
-
thelobbi Bundle Power Bi Fabric AnalyticsDeep expertise in Power BI development including DAX measures, Power Query M transformations, semantic model design, PBIP project scaffolding, REST API workspace management, and Microsoft Fabric integration with Lakehouse and Direct Lake.
-
thelobbi Bundle Fabric Graph And GeoMicrosoft Fabric graph and geospatial analytics - graph model, graph queryset, map, and exploration workflows with preview guardrails.
-
thelobbi Bundle Fabric Data StoreMicrosoft Fabric data store operations - Cosmos DB database, SQL database, Snowflake database links, datamarts, and Event Schema Set governance.
-
thelobbi Bundle Fabric Data ScienceDeep expertise in Microsoft Fabric Data Science — create and track ML experiments with MLflow, train models with scikit-learn/LightGBM/XGBoost/PyTorch in Spark-based notebooks, leverage SynapseML for distributed ML and Cognitive Services, register and version models, batch-score with the T-SQL PREDICT function, and bridge Power BI semantic models to ML workflows via semantic link (SemPy). Targets data scientists and ML engineers building production ML pipelines on Microsoft Fabric.
-
thelobbi Bundle Fabric Data WarehouseDeep expertise in Microsoft Fabric Synapse Data Warehouse — provision warehouses, author T-SQL DDL/DML with auto-distributed storage, design star and snowflake schemas with SCD patterns, load data via COPY INTO and cross-database queries, write stored procedures with error handling, configure row-level and column-level security, monitor query performance with Query Insights, and build dimensional models for the default semantic model. Targets data engineers and analysts working in Fabric workspaces.
-
thelobbi Bundle Power Bi Paginated ReportsThis skill should be used when the user asks about Power BI paginated reports, Report Builder, RDL authoring, or SSRS migration to Fabric. Covers creating pixel-perfect, print-ready reports with tables, matrices, charts, VB.NET expressions, custom code, data source configuration (Fabric Lakehouse, Warehouse, Semantic Model, Dataverse), parameters, subreports, drillthrough, rendering and export (PDF, Excel, Word, CSV), REST API automation, subscriptions, performance tuning, and troubleshooting. Example user requests: "create a paginated invoice report", "write an RDL expression for running totals", "migrate SSRS reports to Fabric", "export a paginated report to PDF via REST API", "fix blank pages in my paginated report".
-
thelobbi Bundle Fabric Data EngineeringDeep expertise in Microsoft Fabric Data Engineering — create and manage lakehouses with OneLake, author PySpark and SparkSQL notebooks, build Delta Lake tables with ACID transactions and time travel, design data pipelines with Copy/Notebook/Dataflow activities, implement medallion architecture (bronze/silver/gold), and optimize Spark workloads for performance. Targets professional data engineers building production Fabric analytics solutions.
-
thelobbi Bundle Fabric Real Time AnalyticsDeep expertise in Microsoft Fabric Real-Time Analytics — create and manage Eventhouses and KQL databases, build eventstreams for streaming ingestion, write KQL (Kusto Query Language) queries with time series analysis and anomaly detection, design Real-Time Dashboards with auto-refresh tiles, configure Data Activator triggers for automated alerting, and integrate with OneLake for unified analytics. Targets data engineers and analysts building production streaming pipelines in Microsoft Fabric.
-
crazymsn Bundle Parallel WebAll-in-one web toolkit powered by parallel-cli, with a strong emphasis on academic and scientific sources. Use this skill whenever the user needs to search the web, fetch/extract URL content, enrich data with web-sourced fields, or run deep research reports. Covers: web search (fast lookups, research, current info — prioritizing peer-reviewed papers, preprints, and scholarly databases), URL extraction (fetching pages, articles, academic PDFs), bulk data enrichment (adding fields to CSV/lists from the web), and deep research (exhaustive multi-source reports grounded in academic literature). Also handles setup, status checks, and result retrieval. Use this skill for ANY web-related task — even if the user doesn't mention 'parallel' or 'web' explicitly. If they want to look something up, fetch a page, enrich a dataset, investigate a topic, find academic papers, check citations, or review scientific literature, this is the skill to use.
-
crazymsn Bundle Optimize For GpuGPU-accelerate Python code using CuPy, Numba CUDA, Warp, cuDF, cuML, cuGraph, KvikIO, cuCIM, cuxfilter, cuVS, cuSpatial, and RAFT. Use whenever the user mentions GPU/CUDA/NVIDIA acceleration, or wants to speed up NumPy, pandas, scikit-learn, scikit-image, NetworkX, GeoPandas, or Faiss workloads. Covers physics simulation, differentiable rendering, mesh ray casting, particle systems (DEM/SPH/fluids), vector/similarity search, GPUDirect Storage file IO, interactive dashboards, geospatial analysis, medical imaging, and sparse eigensolvers. Also use when you see CPU-bound Python code (loops, large arrays, ML pipelines, graph analytics, image processing) that would benefit from GPU acceleration, even if not explicitly requested.
-
crazymsn Bundle Clinical Decision SupportGenerate professional clinical decision support (CDS) documents for pharmaceutical and clinical research settings, including patient cohort analyses (biomarker-stratified with outcomes) and treatment recommendation reports (evidence-based guidelines with decision algorithms). Supports GRADE evidence grading, statistical analysis (hazard ratios, survival curves, waterfall plots), biomarker integration, and regulatory compliance. Outputs publication-ready LaTeX/PDF format optimized for drug development, clinical research, and evidence synthesis.
-
zhizhunbao Bundle Dev WekaWork with Weka data files (.arff format). Use when (1) user needs to convert ARFF to CSV, (2) mentions Weka data files, (3) needs to extract column names from ARFF files.
-
masih-0x3 Bundle Pentest Tools主动渗透测试工具链。覆盖信息收集、端口扫描、漏洞扫描、Web 渗透、SQL 注入、目录爆破、密码破解等场景。 通过 MCP server(pentestMCP / mcp-security-hub)将 20+ 安全工具暴露给 AI agent。 触发关键词:渗透测试、端口扫描、Nmap、漏洞扫描、Nuclei、SQL 注入、SQLMap、目录爆破、FFUF、密码破解、Hashcat、信息收集、子域名、Web 渗透、ZAP、Burp。
-
kevin-liu-01 Skill DBRun read-only SQL queries against Supabase databases (local or remote). Use when inspecting schema, debugging data, checking migrations, or answering questions about database state.
-
kevin-liu-01 Skill DB WriteRun mutating SQL (INSERT, UPDATE, DELETE, DDL) against Supabase databases. Requires explicit user invocation. Use for migrations, seed data, schema changes, or data fixes.
-
naveedharri Bundle Baalda GuideAnswer any question about Baalda (the team second-brain app at baalda.com) in plain, non-technical language — what it is, what it can and cannot do, which file formats it supports (Markdown, images, PDF, DOCX, XLSX, code files), how sync, offline, sharing, permissions, version history, AI/MCP, pricing, platforms and self-hosting work. Use this whenever someone asks "does Baalda…", "can Baalda…", "how does Baalda…", compares it with Obsidian/Notion/Logseq, or asks what happens to a file type in a vault, even if they do not say the word Baalda but are clearly asking about this product's features.
-
naveedharri Bundle Lead GenerationSource, qualify, enrich, and research a B2B lead list from an Ideal Customer Profile, end to end. Use this skill whenever the user wants to "find leads", "source prospects", "build a lead list", "get me leads for [ICP]", "scrape leads", "find companies that match", "build a prospect list", "/lead-gen", or describes who they sell to and wants a contactable, qualified list back. It picks the right data source for the ICP (Google Maps for local businesses, Sales Navigator or LinkedIn scrapers for B2B roles, a prospecting database otherwise), confirms the tools are connected, sources at the right volume, qualifies every lead with parallel subagents, enriches and verifies contact data (email + phone), runs deep per-lead research, and delivers a clean CSV or Google Sheet. Trigger it even when the user does not name a tool, as long as they want leads that match a profile. The output feeds the `outreach` skill.
-
naveedharri Skill Email PersonalizationWrite hyper-personalized cold email icebreakers for B2B leads using their company intelligence and LinkedIn data. Use this skill whenever the user says "write icebreakers", "personalize emails", "email personalization", "cold email first lines", "write opening lines", "personalize outreach", "create email openers", or has enriched leads and wants to write the first line of a cold email for each. Also trigger when the user has a CSV with lead intelligence columns and wants to generate personalized outreach copy. This skill produces icebreakers that sound human, reference real observations, and tie back to the user's product/service.
-
tranhieutt Bundle SQL Optimization PatternsProvides SQL optimization patterns for query performance, indexing strategies, schema design, and database tuning. Use when optimizing slow queries, designing indexes, or tuning database performance.
-
finpeakinc Bundle SnowflakeUse when the user wants to install or verify the official Snowflake CLI, configure or test PAT-authenticated Snowflake connections, execute Snowflake SQL, inspect databases and objects, create or drop Snowflake objects, run SQL files, or operate Snowflake workloads and applications through the `snow` command. Provides an AI-safe Bash wrapper with automatic CLI installation, owner-only PAT token files, JSON output, local-only SQL includes, dry-run previews, and explicit confirmation gates for DDL, DML, connection changes, and destructive operations.
-
finpeakinc Bundle Mysql CrudUse when the user wants to configure saved MySQL database connections, connect directly or through SSH, inspect schemas, query rows, insert records, update records, delete records, or run safe MySQL SQL with a Bash script, dry-run protections, readonly profiles, local profile storage, SSH tunnel access, or remote-server MySQL access through a saved .env DATABASE_URL.
-
finpeakinc Bundle Sqlite CrudUse when the user wants to configure saved SQLite database file profiles, inspect a local SQLite file, list tables or columns, query rows, insert records, update records, delete records, or run safe SQLite SQL with a Bash script, sqlite3, dry-run protections, readonly profiles, and local profile storage for reusable local file paths.
-
finpeakinc Bundle Postgresql CrudUse when the user wants to configure saved PostgreSQL database connections, connect directly or through SSH, inspect schemas, query rows, insert records, update records, delete records, or run safe PostgreSQL SQL with a Bash script, dry-run protections, readonly profiles, local profile storage, SSH tunnel access, or remote-server PostgreSQL access through a saved .env DATABASE_URL.
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include media-scenedetect, tauri-plugins, doris-docs. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.