Text Summarization Eval

Evaluates the abstractive text summarization capability of large language models by measuring how well they condense news articles into coherent, factually faithful, and linguistically natural summaries compared to human-written references. Use when the user wants to benchmark on CNN/Daily Mail 3.0.0, XSum, or asks about evaluating this task. Reports BERT Score.

qhjqhj00 a7b235d 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/text-summarization-eval commit a7b235d068

Frequently asked questions

npx skillmds add qhjqhj00/text-summarization-eval