Tricky2 Benchmark Evaluating Human Error

Taxonomy-guided analysis of mixed human+LLM bugs in code. Classifies bug origins, localizes interacting defects, and repairs hybrid-origin errors. Use when: 'review this AI-generated code for bugs', 'find bugs in this human+AI codebase', 'classify whether this bug is human or LLM', 'audit code written by both humans and copilot', 'debug interacting errors in mixed-origin code', 'analyze bug patterns in AI-assisted development'.

ndpvt-web 842fa2c 15.1 KB Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/tricky2-benchmark-evaluating-human-error commit 842fa2c155

Frequently asked questions

npx skillmds@latest add ndpvt-web/tricky2-benchmark-evaluating-human-error