SkillMD SkillMD
Skills
All skills Official skills Leaderboard Saved
Categories
Coding & Dev Tools9772AI & ML5010DevOps & Infra2864Integrations & APIs2327Productivity2190Security1976 All categories →
Packs Docs
menu-rounded
Skills Categories Packs Docs My skills Saved
light-dark-mode
Profile My skills Saved Collections Preferences Submit a skill
All Skills 25,790 ✦ Verified
Categories
AI & ML 5,010
Agent Building 585 Image & Video Generation 189 MCP Servers 94 Model Training & Fine-tuning 377 Prompt Engineering 130 RAG & Embeddings 120 Speech & Audio 107
Coding & Dev Tools 9,772 Data & Analytics 1,725 Design & Media 983 DevOps & Infra 2,864 Docs & Writing 1,294 Finance & Business 347 Integrations & APIs 2,327 Marketing & Growth 1,430 Product & Planning 1,037 Productivity 2,190 Research & Search 928 Security 1,976 Web & Frontend 1,639

Results for “summarization”

6 skills
qhjqhj00
ndcg-10
Evaluates how well internal model representations (hidden states) predict token-level information importance in summarization tasks, using NDCG@10 and Spearman's rank correlation.
3
qhjqhj00
menli
Evaluates the robustness and alignment with human judgment of reference-based and reference-free evaluation metrics for machine translation and summarization, particularly under adversarial conditions.
3
qhjqhj00
reflex
Evaluates machine-generated log summaries without human-written references, using LLM judgment and dense embeddings to score relevance, informativeness, and coherence.
3
More results
qhjqhj00
apessrc
Evaluates the faithfulness of abstractive summaries by verifying if factual claims (masked as cloze questions) in the reference summary can be correctly answered using only the generated summary, compared against a gold-standard answer derived from the source context.
3
qhjqhj00
l-eval
Benchmarks long-context language models across 20 sub-tasks spanning 3k–200k tokens, covering retrieval, reasoning, summarization, and instruction understanding, with exact-match accuracy as the primary metric.
3
qhjqhj00
feqa
Evaluates the faithfulness of abstractive summaries by generating questions from summary sentences and verifying if the answers can be extracted from the source document, reporting Pearson and Spearman correlations with human judgments.
3
SKILLMD.com

The open registry of AI Agent Skills: safety-reviewed SKILL.md files for Claude, Cursor, Codex & 60+ agents.

$ npm i skillmds

Explore

All Skills Categories Agents Packs New & Latest Leaderboard

Support

About Contact npm Terms Privacy

Learn

Docs Blog Stats FAQ Submit a Skill
© 2026 SkillMD.com Skills attributed to their authors under their original licenses.
SKILLMD