Lexsumm Eval

This benchmark evaluates the ability of sequence-to-sequence models to generate accurate, concise, and faithful summaries of long legal documents across multiple jurisdictions. It probes domain-specific summarization capabilities, testing how well models handle varying input lengths, compression ratios, and the balance between extractive and abstractive generation in legal English. Use when the user wants to benchmark on BillSum, EurLexSum, GovReport, MultiLexSum-Long, MultiLexSum-Short, MultiLexSum-Tiny, InAbs, UKAbs, or asks about evaluating this task. Reports Compression Ratio.

qhjqhj00 13c6640 4.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/lexsumm-eval commit 13c664015f

Frequently asked questions

npx skillmds add qhjqhj00/lexsumm-eval