Long Form Scientific Summarization Eval

Evaluates long-form scientific summarization models on their ability to generate relevant and faithful abstracts across clinical, chemical, and biomedical domains. It probes how calibration set construction and candidate selection strategies affect model performance on standard relevance and faithfulness metrics. Use when the user wants to benchmark on Scientific Summarization Datasets, or asks about evaluating this task. Reports Rouge-1 F1.

qhjqhj00 56e327c 4.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/long-form-scientific-summarization-eval commit 56e327c649

Frequently asked questions

npx skillmds add qhjqhj00/long-form-scientific-summarization-eval