Eval Artifacts

LLM-as-judge scoring skill. Takes a transcript + sandbox state, applies the rubric, and produces a structured score report in docs/testingResults/.

NoahJenkins Updated

File contents

NoahJenkins/Copilot-Stuff/tree/main/.github/skills/eval-artifacts commit 73fe1f5fcd

Frequently asked questions

npx skillmds@latest add noahjenkins/eval-artifacts