Data & Analytics
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
-
undermybelt Bundle Hstrn Server Ot EnvrnmThis skill covers hardening and securing process historian servers (OSIsoft PI, Honeywell PHD, GE Proficy, AVEVA Historian) in OT environments. It addresses network placement across Purdue levels, access control for historian interfaces, data replication through DMZ using data diodes or PI-to-PI connectors, SQL injection prevention in historian queries, and integrity protection of process data used for safety analysis, regulatory reporting, and process optimization.
-
findscripter Skill Cim Builder当为卖方并购流程(sell-side)准备募售材料、起草保密信息备忘录、把公司资料整理成投资人可读文档时使用;产出含执行摘要/公司/行业/增长/客户/运营/财务/附录的 40-60 页 .docx 及 Excel 财务附录;不适用于买方材料、法律文书起草或正式估值建模。触发词:CIM、保密信息备忘录、募售材料、卖方材料、info memo、offering memorandum
-
findscripter Skill CSV Data Cleaner当需要清洗 CSV/表格数据——去重、缺失值处理、类型规整、异常值、列标准化时使用;触发词:数据清洗、去重、缺失值、脏数据、规整。
-
findscripter Skill Dataset Profiler当拿到陌生表/文件、动手分析前要先摸清其形状、质量与规律时使用;做全表+逐列画像(行列/粒度/主键、空值率、基数、分布、Top/Bottom 值),分类列角色、标记质量隐患,产出画像摘要表+问题清单+可跟进分析建议;不适用于深度修复清洗、设计 schema、搭 ETL 管道。触发词:数据画像、探查数据、profiling、空值率、基数分布、新表先看什么、该用哪些维度指标
-
nota-america Skill Recipe Create Expense TrackerSet up a Google Sheets spreadsheet for tracking expenses with headers and initial entries.
-
nota-america Skill Recipe Sync Contacts To SheetExport Google Contacts directory to a Google Sheets spreadsheet.
-
nota-america Skill Create VizCreate publication-quality visualizations with Python. Use when turning query results or a DataFrame into a chart, selecting the right chart type for a trend or comparison, generating a plot for a report or presentation, or needing an interactive chart with hover and zoom.
-
nota-america Skill Recipe Create Events From SheetRead event data from a Google Sheets spreadsheet and create Google Calendar entries for each row.
-
nota-america Skill Write QueryWrite optimized SQL for your dialect with best practices. Use when translating a natural-language data need into SQL, building a multi-CTE query with joins and aggregations, optimizing a query against a large partitioned table, or getting dialect-specific syntax for Snowflake, BigQuery, Postgres, etc.
-
nota-america Skill SQL QueriesWrite correct, performant SQL across all major data warehouse dialects (Snowflake, BigQuery, Databricks, PostgreSQL, etc.). Use when writing queries, optimizing slow SQL, translating between dialects, or building complex analytical queries with CTEs, window functions, or aggregations.
-
nota-america Bundle Ray DataScalable data processing for ML workloads. Streaming execution across CPU/GPU, supports Parquet/CSV/JSON/images. Integrates with Ray Train, PyTorch, TensorFlow. Scales from single machine to 100s of nodes. Use for batch inference, data preprocessing, multi-modal data loading, or distributed ETL pipelines.
-
findscripter Skill Dbt Transformation Modeler当用 dbt 搭建分析工程数据转换管道、按 staging/intermediate/marts 分层建模、加数据质量测试或做增量处理时使用;产出分层模型结构、SQL 模型、sources/schema YAML、测试与增量策略配置;不适用于实时流处理、非 dbt 的 ETL 编排或纯 SQL 即席查询;触发词:dbt、data build tool、分析工程、analytics engineering、数据建模、staging/marts 分层、medallion、增量模型、incremental、dbt 测试、source freshness、维度事实表 dim/fct。
-
findscripter Skill Seaborn Statistical Charts当需要用 Seaborn 从 DataFrame 直接画统计图(散点/折线/分布/箱线/小提琴/热力图/回归/分面网格)并要出版级美观默认时使用;做选图型→建图→分面/语义映射→调主题配色→存图的可执行流程,产出图像文件;不适用于交互式/Web 图表(Plotly/Bokeh)、纯 matplotlib 底层绘制、地理/网络专用图。触发词:seaborn、统计图、热力图、分布图、分面、出版级配图
-
findscripter Skill Smb Cash Flow Forecast当小微企业主问"能否发出工资/还剩多少跑道/会不会现金断流"、要做 30/60/90 天现金流预测时使用;从 QuickBooks/PayPal/Stripe/Square 或 CSV 读取 AR/AP 与固定成本,按各客户历史回款时滞与方差算预期流入流出、带置信区间,输出聊天摘要+可下载 XLSX 并点名风险;不适用于正式做账报税、财报审计或上市公司估值建模。触发词:现金流预测、能不能发工资、现金跑道、现金断流、cash crunch、runway
-
findscripter Skill Backend Security Coder当编写或评审后端代码与 API、需要防注入/认证授权/安全响应时使用;做输入校验、参数化查询、JWT/会话、CSRF/SSRF 防护、安全响应头与限流的落地实现与加固清单;不适用于纯前端、合规审计/威胁建模/渗透测试规划(交 security-auditor)。触发词:SQL注入、JWT、CSRF、限流、安全响应头
-
findscripter Skill Autodock Vina Docking当已知结合口袋、需预测小分子与蛋白靶点的结合构象/亲和力或做虚拟筛选打分排序时使用;用 Vina Python API + Meeko/RDKit 完成受体与配体备制、定义搜索盒、对接打分、构象与结合能分析及批量筛选,产出 PDBQT 构象与能量排名 CSV;不适用于结合位点未知的盲对接(用 DiffDock)、需要 CNN 打分(用 GNINA)或自由能微扰类高精度计算。触发词:分子对接、docking、AutoDock Vina、PDBQT、虚拟筛选、virtual screening、结合能、binding affinity、Meeko、对接盒。
-
findscripter Skill Flowio Flow Cytometry当需要用 Python 读写流式细胞 FCS 文件(v2.0–3.1)、把事件数据取成 NumPy 数组、提取通道元数据、转 DataFrame/CSV、生成或批处理 FCS 时使用;产出事件矩阵、通道名/范围、CSV 与新 FCS 文件;不适用于补偿(compensation)、设门(gating)、FlowJo 工作区(改用 FlowKit)或散点/密度可视化(改用 matplotlib);触发词:FCS、flow cytometry、流式细胞、FlowIO、FlowData、as_array、pnn_labels、create_fcs、多数据集 FCS。
-
findscripter Skill Rdkit Cheminformatics当需要用 RDKit 跑「化合物库画像 + 虚拟筛选」端到端流程时使用;批量解析 SMILES/SDF、标准化去重、算描述符、Lipinski/Veber 类药性过滤、Morgan 指纹 Tanimoto 相似度筛选、SMARTS 子结构过滤、Butina 聚类、反应枚举、2D/3D 构象与绘图,产出描述符表、命中 SDF/CSV 与分子图;不适用于追求更简接口(用 datamol)、蛋白对接/分子动力学/量子化学、纯数据库检索;触发词:rdkit、虚拟筛选、化合物库、描述符、Lipinski、Tanimoto、Morgan 指纹、SMARTS、Butina 聚类、构象生成
-
findscripter Skill Scikit Image Bioimage当用 Python 处理显微/荧光生物图像(TIFF/PNG,NumPy 数组)需做读写、滤波去噪、阈值/分水岭分割、形态学、区域属性测量、斑点/特征检测时使用;用 scikit-image+SciPy 跑分割-测量流程并产出标注掩膜、regionprops 表格(CSV)与叠加图;不适用于实时视频(用 OpenCV)、深度学习触碰细胞分割(用 CellPose)、交互式多维可视化(用 napari);触发词:scikit-image、skimage、显微图像、细胞核分割、watershed、regionprops、阈值、形态学、blob 检测
-
findscripter Skill Three Statement Model当需要在已有 Excel/openpyxl 模板中填充并联动利润表、资产负债表、现金流量表时使用;做法是用公式(而非硬编码值)填充历史数与假设驱动、建立三表勾稽并逐表校验产出可审计的三表模型;不适用于从零设计模型版式、估值/DCF/LBO 建模或纯数据录入。触发词:三表模型、3-statement model、财务建模、利润表/资产负债表/现金流量表联动、IS BS CF、资产负债表平衡、现金勾稽、cash tie-out、情景分析
-
findscripter Skill Dask Distributed Dataframes当 pandas/NumPy 工作流超出内存或需跨核/跨机并行(约 100GiB 单机到 100TiB 集群)时使用;用 Dask 的 DataFrame/Array/Bag/Futures 构建惰性任务图并行执行,产出聚合结果或 Parquet/Zarr 落地;不适用于内存可容纳追求极速(用 polars)或单机核外分析(用 vaex)。触发词:dask、超内存、larger-than-RAM、并行、分布式、map_partitions
-
findscripter Skill Data Throughput Accelerator当大规模数据导入/回填/导出/ETL/数仓装载/清单追平/表同步需要在保证正确性的前提下显著提速时使用;做瓶颈分层定位与多变体基准对比,产出最快且行数/时间戳一致的可固化路径与硬核对账块;不适用于小数据量、单纯调度编排或与正确性无关的纯算力问题;触发词:回填、ETL提速、数仓装载
-
findscripter Skill Dbt Transformation Patterns当用 dbt 在数据仓库上搭建分层转换管道(staging/intermediate/marts)、加测试与文档、做增量模型时使用;产出分层命名规范、source/staging/mart 模型与 schema.yml 测试、增量物化策略及常用 dbt 命令清单;不适用于无 dbt/仓库的纯即席 SQL 查询或无源数据访问权限的场景。触发词:dbt、数据建模、增量模型
-
findscripter Skill Spreadsheet Formula Auditor当需要审计 Excel/电子表格公式正确性或核查财务模型(DCF/LBO/三表)勾稽完整性时使用;按选区/单张表/整本模型逐层排查公式错误、硬编码、勾稽断点并产出分级问题清单;不适用于纯数据清洗、图表制作或无公式的静态表。触发词:审计表格、检查公式、模型核对、对不上、勾稽、audit spreadsheet、check formulas、model won't balance
-
findscripter Skill People Analytics Report当需要为管理层拉编制快照/分析团队流失趋势/准备多元化代表性指标/评估管理幅度与离职风险时使用;从 HR 数据(CSV 或 HRIS)做人力分析,产出含执行摘要+关键指标表+建议+口径说明的报表;不适用于真实抓取 HRIS 数据、个人发薪/HRIS 写操作、绩效打分与合规裁定;触发词:人力报表、编制快照、流失分析、多元化、组织健康度、管理幅度、离职风险、headcount、attrition
-
findscripter Skill Sector Landscape Report当需要为客户、行业首次覆盖或主题研究撰写行业/赛道格局综述时使用;做市场规模—竞争格局—估值—投资含义全景调研并产出综述报告(Word/PPT+Excel附录);不适用于单家公司深度估值建模或个股买卖建议;触发词:行业报告、赛道综述、市场格局、主题研究
-
findscripter Skill Astronomy Data Toolkit当用 Python 处理天文/天体物理数据(坐标变换、单位换算、FITS 读写、表格、时间系统、WCS、宇宙学)时使用;用 Astropy 编写或调试分析代码并产出量纲一致、可复现的结果;不适用于通用数据清洗或非天文数值计算(改用 pandas/numpy)。触发词:astropy、天文数据、astronomy、SkyCoord、坐标变换、FITS、WCS、宇宙学、cosmology、天文单位、Quantity
-
findscripter Skill Gget Genomic Databases当需要用一个统一的 Python/CLI 接口(gget)跨 Ensembl、UniProt、NCBI、BLAST/BLAT、AlphaFold、Enrichr、OpenTargets、CELLxGENE、cBioPortal/COSMIC、ARCHS4 等 20+ 基因组数据库做基因查询、取序列、比对、结构预测、富集与疾病/药物关联时使用;做选模块、调 gget 取数并产出 DataFrame/JSON/FASTA/PDB 结果;不适用于大批量或高级 BLAST 参数(用 biopython)与带限速的多库 SDK 编排(用 bioservices);触发词:gget、Ensembl 查基因、gget search/info/seq、BLAST、AlphaFold、Enrichr、OpenTargets、CELLxGENE、cBioPortal
-
findscripter Skill Odoo Performance Tuner当 Odoo 生产环境变慢、超时、报 MemoryError/Worker timeout,或需为指定服务器规格调优 odoo.conf 时使用;做 worker/内存/超时参数计算、用 pg_stat_statements 定位慢 SQL 与缺失索引、用内置 Profiler 抓 Python+SQL 链路并输出可落地配置改动;不适用于 Odoo 业务建模/二次开发、深度 PostgreSQL 参数(shared_buffers 等用 PGTune)、前端 JS 渲染性能、Odoo.sh 受限托管的底层调参;触发词:odoo 慢、worker timeout、MemoryError、odoo.conf 调优、慢查询、pg_stat_statements、缺失索引、odoo profiler、N+1
-
findscripter Skill Odoo XML Views Builder当为 Odoo 模型编写或修复 Form/List/Kanban/Search/Calendar/Graph 视图 XML 时使用;做生成可直接粘贴的 ir.ui.view 视图定义并处理 v14-16 的 attrs 到 v17 内联表达式迁移、visibility/groups/domain/widget 配置;不适用于 OWL/JS 组件、searchpanel、website QWeb 模板及企业版 Cohort/Map 视图;触发词:odoo 视图、ir.ui.view、form/list/kanban/search 视图、attrs、invisible、statusbar、notebook
-
findscripter Skill Pe Returns Sensitivity当评估私募股权(PE/Buyout)交易、压力测试假设或准备投委会(IC)回报材料时使用;据交易/融资/经营/退出假设快速搭建 IRR/MOIC 基准回报、二维敏感性表、三档情景与回报归因,产出 Excel 与一页 IC 摘要;不适用于公开股权估值、DCF 企业估值或非杠杆收购建模;触发词:returns analysis、IRR sensitivity、MOIC table、回报分析、敏感性表、估算回报、back of the envelope
-
findscripter Skill Portfolio Risk Metrics当需要度量组合风险、设置风险限额或搭建风险监控/报表时使用;用 Python(numpy/pandas/scipy) 从收益率序列计算 VaR、CVaR、夏普、索提诺、卡玛、最大回撤、Beta、风险平价等并产出风险摘要;不适用于价格抓取、回测撮合、择时/选股信号生成与因子归因建模;触发词:风险指标、VaR、CVaR、夏普比率、Sortino、最大回撤、risk metrics、drawdown、风险平价
-
findscripter Skill Apify Ecommerce Scraper当需要从亚马逊、沃尔玛等电商平台批量抓取商品、价格、库存、评论或卖家数据(用于比价、MAP 监控、竞品分析、评论情感/质量分析、卖家发现)时使用;调用 Apify「e-commerce-scraping-tool」Actor,按三类工作流配置输入并导出 CSV/JSON 结果与洞察;不适用于无 APIFY_TOKEN、需自建爬虫绕过反爬、或抓取非电商页面。触发词:电商抓取、商品比价、价格监控、评论分析、卖家发现、apify、ecommerce scraping、product price、review scraping
-
findscripter Skill Apify Multi Platform Scraper当需要从社媒/地图/搜索等平台抓公开数据但还没选定具体 Apify Actor 时使用;做 AI 选 Actor→取 schema→运行→导出 CSV/JSON 的统一抓取流水线,覆盖 55+ Actor;不适用于登录/付费墙数据、无 APIFY_TOKEN、纯本地解析或写自定义爬虫。触发词:Apify、抓 Instagram/TikTok/YouTube/Facebook、Google Maps 商家、爬社媒、线索采集
-
nota-america Bundle SpreadsheetUse when tasks involve creating, editing, analyzing, or formatting spreadsheets (`.xlsx`, `.csv`, `.tsv`) with formula-aware workflows, cached recalculation, and visual review.
-
findscripter Skill Browser Automation Builder当需要用 Playwright 抓取网页数据、自动填表登录、截图存档、提取结构化数据或搭建可重复浏览器工作流时使用;产出含选择器策略、分页/会话/反检测/重试模式的 Python 自动化脚本与 JSON/CSV/JSONL 数据;不适用于写 E2E 测试(用 playwright-pro)、纯 API 测试或性能压测;触发词:网页抓取、自动填表、截图存PDF、反爬反检测、SPA动态内容、会话复用
Frequently asked questions
What are Data & Analytics agent skills?
Data agent skills make AI agents useful for data work: writing SQL, cleaning datasets, building pipelines, working with spreadsheets, and producing analyses. Each skill is a reviewed SKILL.md file that teaches the agent one workflow well, ready to install in seconds.
Which Data & Analytics skills are most installed?
Popular Data & Analytics skills on SkillMD right now include hstrn-server-ot-envrnm, cim-builder, csv-data-cleaner. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do Data & Analytics skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.