Figedit Chart Editing Eval

This benchmark evaluates the ability of vision-language and image-editing models to perform semantically correct, structure-aware modifications to scientific charts. It probes whether models can follow precise editing instructions while preserving data-encoding consistency, axis coherence, and legend integrity, rather than merely producing pixel-level visual similarity. Use when the user wants to benchmark on FigEdit, or asks about evaluating this task. Reports Instruction-following score.

qhjqhj00 a5ce1c1 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/figedit-chart-editing-eval commit a5ce1c119f

Frequently asked questions

npx skillmds add qhjqhj00/figedit-chart-editing-eval