SkillMD SkillMD
Skills
All skills Official skills Leaderboard Saved
Categories
Coding & Dev Tools9777AI & ML5020DevOps & Infra2868Integrations & APIs2327Productivity2192Security1976 All categories →
Plugins Docs
menu-rounded
Skills Categories Plugins Docs My skills Saved
light-dark-mode
Profile My skills Saved Collections Edit profile Submit a skill
All Skills 25,836 ✦ Verified
Categories
AI & ML 5,020 Coding & Dev Tools 9,777 Data & Analytics 1,729
Data Analysis 671 Data Visualization 86 ETL & Pipelines 153 SQL & Databases 340 Spreadsheets 27
Design & Media 985 DevOps & Infra 2,868 Docs & Writing 1,311 Finance & Business 347 Integrations & APIs 2,327 Marketing & Growth 1,430 Product & Planning 1,037 Productivity 2,192 Research & Search 928 Security 1,976 Web & Frontend 1,641

Results for “metric-evaluation”

6 skills
nvidia
RAG Eval
Evaluates RAG pipelines using a filesystem-based benchmark with corpus/ and train.json, running evaluate_rag.py to tune retrieval and generation flags and interpret RAGAS metrics.
2.2k · bundle
More results
k-dense-ai
Pytdc
Access AI-ready drug discovery datasets and benchmarks from Therapeutics Data Commons, covering ADME, toxicity, drug-target interactions, and molecular generation with standardized splits and evaluation metrics.
30.2k · bundle
wondelai
Lean Analytics
Choose and audit startup metrics using the Lean Analytics framework: separate actionable metrics from vanity metrics, identify the One Metric That Matters for your business model and stage, set targets, and plan instrumentation.
1.6k · bundle
agricidaniel
Ads Test
Design and evaluate paid-ad experiments with hypotheses, randomization, sample-size calculations, guardrails, and decision rules for A/B and split tests.
nexu-io
Experiment Readout
Transforms A/B test and product experiment data into actionable readouts with hypothesis, metrics, interpretation, and decision.
· bundle
k-dense-ai
Scikit Survival
Perform survival analysis and time-to-event modeling in Python using scikit-survival, including Cox models, random survival forests, gradient boosting, survival SVMs, and evaluation metrics like concordance index and Brier score.
30.2k · bundle
SKILLMD.com

The open registry of AI Agent Skills: safety-reviewed SKILL.md files for Claude, Cursor, Codex & 60+ agents.

$ npm i skillmds

Explore

All Skills Categories Agents Plugins New & Latest Leaderboard

Support

About Contact npm Terms Privacy

Learn

Docs Blog Stats FAQ Submit a Skill
© 2026 SkillMD.com Skills attributed to their authors under their original licenses.
SKILLMD