Orangesum Eval

Evaluates abstractive French summarization quality by measuring lexical overlap, semantic similarity, and human judgments of accuracy, informativeness, and fluency. Use when the user wants to benchmark on OrangeSum, or asks about evaluating this task. Reports ROUGE-L.

qhjqhj00 a30bb45 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/orangesum-eval commit a30bb454b4

Frequently asked questions

npx skillmds add qhjqhj00/orangesum-eval