Wcep Mds Eval

Evaluates multi-document summarization systems on news event clusters by measuring how well generated summaries match human-written reference summaries. It probes the model's ability to extract or generate concise, informative summaries from highly redundant, large-scale document collections. Use when the user wants to benchmark on WCEP, or asks about evaluating this task. Reports ROUGE F1-score.

qhjqhj00 5db7203 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/wcep-mds-eval commit 5db7203786

Frequently asked questions

npx skillmds add qhjqhj00/wcep-mds-eval