Fus Multimodal Robot Eval

Evaluates a robot policy's ability to ground heterogeneous sensor modalities (vision, touch, sound) into language instructions for zero-shot task execution in partially observable environments. It probes multimodal prompting, compositional reasoning, and the necessity of auxiliary contrastive and language grounding losses. Use when the user wants to benchmark on WidowX Multimodal Teleoperation Dataset, or asks about evaluating this task. Reports task success.

qhjqhj00 f0e5a80 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/fus-multimodal-robot-eval commit f0e5a80df1

Frequently asked questions

npx skillmds add qhjqhj00/fus-multimodal-robot-eval