Noticia Eval

This benchmark evaluates large language models' ability to interpret misleading clickbait headlines and extract the core information buried in Spanish news articles. It probes the models' capacity for ultra-concise abstractive summarization in a multilingual setting, specifically testing whether they can ignore irrelevant article content and produce brief, accurate summaries. Use when the user wants to benchmark on NoticIA, or asks about evaluating this task. Reports ROUGE-1.

qhjqhj00 c1fb28a 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/noticia-eval commit c1fb28a1fa

Frequently asked questions

npx skillmds add qhjqhj00/noticia-eval