Cfbenchmark Mm Eval

Evaluates multimodal large language models' ability to interpret financial charts, tables, and diagrams in Chinese, and answer domain-specific questions. It probes visual reasoning, statistical and structural analysis, and financial concept comprehension under zero-shot conditions. Use when the user wants to benchmark on CFBenchmark-MM, or asks about evaluating this task. Reports accuracy.

qhjqhj00 05281c4 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cfbenchmark-mm-eval commit 05281c488c

Frequently asked questions

npx skillmds add qhjqhj00/cfbenchmark-mm-eval