Overcooked AI Adaptation Eval

This benchmark evaluates the real-time adaptability and communication capabilities of LLM-powered embodied agents in human-robot collaboration. It probes how well agents adjust their high-level subtask planning and low-level movement paths when faced with dynamic, constrained environments and non-adaptive human partners. Use when the user wants to benchmark on Enhanced Overcooked-AI, or asks about evaluating this task. Reports overall score.

qhjqhj00 88a6c1e 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/overcooked-ai-adaptation-eval commit 88a6c1e291

Frequently asked questions

npx skillmds add qhjqhj00/overcooked-ai-adaptation-eval