Structured Output Benchmark Eval

Evaluates large language models' ability to extract structured information from multi-modal sources (text, images, audio) into valid JSON formats, isolating schema compliance from value accuracy. Use when the user wants to benchmark on Multi-Source Structured Output Benchmark, or asks about evaluating this task. Reports correct_value_extraction.

qhjqhj00 873bd26 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/structured-output-benchmark-eval commit 873bd26b6f

Frequently asked questions

npx skillmds add qhjqhj00/structured-output-benchmark-eval