Neural Mmo Eval

This benchmark evaluates the robustness and generalization of multi-agent reinforcement learning policies in a large-scale, open-ended simulation. It probes a model's ability to cooperate with teammates and compete against unknown opponents or fixed baselines across varying difficulty levels and dynamic environments. Use when the user wants to benchmark on Neural MMO, or asks about evaluating this task. Reports TrueSkill.

qhjqhj00 2cc29ed 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/neural-mmo-eval commit 2cc29ed4ab

Frequently asked questions

npx skillmds add qhjqhj00/neural-mmo-eval