Clawpathy Autoresearch

Eval-driven skill tuning. Given a task and an LLM-judge rubric, iteratively rewrites a SKILL.md until a downstream executor agent performs well against the judge. Low-code: all evaluation is LLM-as-judge, not deterministic Python.

stanfish06 Updated

File contents

stanfish06/skillquarium/tree/main/skills/clawpathy-autoresearch commit 7f5e329e23

Frequently asked questions

npx skillmds@latest add stanfish06/clawpathy-autoresearch