Cawplan Internal QA Coding Humaninputs Test

Internal QA check for the AI-coding human-input classifier: pulls already-uploaded human inputs via the cawplan CLI, classifies each one via cawplan-internal-qa-coding-humaninputs, and compares category and/or topic against persisted cloud labels — in category-only, topic-only, or both-together mode — reporting accuracy and concrete mismatches for manual review. Use when: asked to test/verify/check human-input classification accuracy — e.g. "test today's category accuracy", "check topic only for spx last week", "compare both category and topic for the last 2 days" — optionally scoped to one person, one product, or both. NOT for: classifying a single ad hoc sentence (use cawplan-internal-qa-coding-humaninputs directly), submitting coding reports (use cawplan-coding-commit), general cost/usage insights or prompt-quality scores (use cawplan-coding-insights), or creating tickets.

cawcut a286a18 9.5 KB Updated

File contents

cawcut/skill-cawplan/tree/main/skills/cawplan-internal-qa-coding-humaninputs-test commit a286a18d70

Frequently asked questions

npx skillmds@latest add cawcut/cawplan-internal-qa-coding-humaninputs-test