Alfred Eval

Embodied instruction following in a simulated household environment, requiring an agent to execute long-horizon navigation and object manipulation tasks based on natural language commands. Use when the user wants to benchmark on ALFRED, or asks about evaluating this task. Reports Success Rate (SR).

qhjqhj00 4713f77 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/alfred-eval commit 4713f7751f

Frequently asked questions

npx skillmds add qhjqhj00/alfred-eval