Embodied Eval Automation

Plan, build, run, monitor, validate, transfer, and audit reproducible batch episode collection for policy, VLA, world-model, or hybrid models on embodied benchmarks and simulators. Use when a user provides or wants to connect a local/SSH/cloud GPU host, reuse existing model or benchmark assets, obtain GitHub or Hugging Face resources, compare native and unified episode formats, create adapters and validators, stage single-request through batch runs, recover interrupted jobs, manage disk and GPU limits, or deliver auditable episode datasets and reports.

Evidentiary-operatingmicroscope154 26a9618 23 files · 64.6 KB Updated

File contents

Evidentiary-operatingmicroscope154/embodied-eval-automation/tree/main/skills/embodied-eval-automation commit 26a9618da4

Frequently asked questions

npx skillmds@latest add evidentiary-operatingmicroscope154/embodied-eval-automation