Mlsum Eval

Evaluates abstractive text summarization models across multiple languages (French, German, Spanish, Russian, Turkish) to measure generation quality and investigate cross-lingual performance gaps and model biases. Use when the user wants to benchmark on MLSUM, or asks about evaluating this task. Reports ROUGE-L.

qhjqhj00 a4d7e63 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mlsum-eval commit a4d7e63598

Frequently asked questions

npx skillmds add qhjqhj00/mlsum-eval