Tooluniverse Self Review

Review existing work against the user's actual goal and surface evidence-backed strengths, gaps, risks, and next fixes. Use when asked to eval, evaluate, review, assess, or check current/this/my/our work; decide whether a task is complete; build a definition-of-done checklist or rubric; or perform grading, LLM-as-judge, Qworld, or RET evaluation. Treat plain eval/review requests as qualitative: resolve "current work" from the conversation, artifacts, files, or diff, and never assign numeric scores unless the user explicitly requests scores, grades, points, ratings, weighted criteria, Qworld, or RET. Do not use for implementing automated eval suites, tests, graders, or benchmarks.

mims-harvard Updated

File contents

mims-harvard/tooluniverse/tree/main/skills/tooluniverse-self-review commit 1ba7014a15

Frequently asked questions

npx skillmds@latest add mims-harvard/tooluniverse-self-review