Cnndm Eval

Evaluates abstractive summarization quality by scoring generated summaries against human-written references and expert rubric-based scores. It measures how well automatic metrics correlate with human judgments across different summarization systems. Use when the user wants to benchmark on CNNDM, or asks about evaluating this task. Reports COMET.

qhjqhj00 5138126 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cnndm-eval commit 51381263c9

Frequently asked questions

npx skillmds add qhjqhj00/cnndm-eval