Human Eval Review

Generate a single self-contained HTML file for human-in-the-loop evaluation of model output. The reviewer opens it in a browser, makes per-item judgments (pick the matching candidate(s), choose one, rate on a scale, yes/no, or free-text correction), and exports their decisions as JSON. Use when you have analysis output — candidate matches, rankings, classifications, translations, extractions — that a human needs to review, confirm, score, or correct.

kltng Updated

File contents

kltng/humanities-skills/tree/main/human-eval-review commit a87882b07b

Frequently asked questions

npx skillmds@latest add kltng/human-eval-review