Factual Scene Graph Parsing Eval

This benchmark evaluates a model's ability to parse natural language captions into structured scene graphs that faithfully represent described visual elements. It probes compositional generalization and output consistency by testing parsers on both standard and length-constrained splits, measuring how well generated graph structures align with human-annotated ground truth. Use when the user wants to benchmark on FACTUAL, or asks about evaluating this task. Reports SPICE.

qhjqhj00 7650389 4.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/factual-scene-graph-parsing-eval commit 7650389b4a

Frequently asked questions

npx skillmds add qhjqhj00/factual-scene-graph-parsing-eval