Bee 8b Mllm Eval

Evaluates the visual reasoning, factual accuracy, OCR, chart understanding, and mathematical capabilities of fully open multimodal large language models (MLLMs) against a comprehensive suite of established benchmarks. The protocol tests the model's ability to process images and text prompts, generate responses in a thinking mode, and achieve high scores across general VQA, document/chart analysis, and complex math/reasoning tasks. Use when the user wants to benchmark on Bee-8B Evaluation Benchmarks, or asks about evaluating this task. Reports accuracy.

qhjqhj00 d7da517 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/bee-8b-mllm-eval commit d7da517c92

Frequently asked questions

npx skillmds add qhjqhj00/bee-8b-mllm-eval