Long Doc Rouge Eval

Evaluates the quality of abstractive summaries for long scientific documents by measuring n-gram overlap between generated text and reference abstracts. Use when the user wants to benchmark on arXiv, PubMed, or asks about evaluating this task. Reports ROUGE-1.

qhjqhj00 f63277c 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/long-doc-rouge-eval commit f63277c961

Frequently asked questions

npx skillmds add qhjqhj00/long-doc-rouge-eval