Eval

Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winner among completed AgentHub agents.

nous-hermeshub 1f1beb4 2.4 KB Updated 1 repo stars

File contents

nous-hermeshub/hermes-community-hub/tree/main/skills/productivity/eval commit 1f1beb4f64

Frequently asked questions

npx skillmds add nous-hermeshub/eval