Jobresqa Benchmark Machine Reading

Build and evaluate multilingual machine reading comprehension systems for HR documents (resumes and job descriptions). Implements the JobResQA pipeline: tiered QA generation, cross-document reasoning, placeholder-based bias testing, TEaR translation, and G-Eval LLM-as-judge scoring. Use when: 'evaluate resume parsing accuracy', 'build HR question answering', 'test multilingual resume understanding', 'check bias in resume screening', 'cross-document QA on resumes and JDs', 'benchmark LLM on HR tasks'.

ndpvt-web 3b714f1 16.3 KB Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/jobresqa-benchmark-machine-reading commit 3b714f119a

Frequently asked questions

npx skillmds@latest add ndpvt-web/jobresqa-benchmark-machine-reading