Qrcd Eval

Evaluates machine reading comprehension on a low-resource religious domain (Qur'an). It probes a model's ability to extract precise answer spans from Arabic text given a question, testing both exact matching and partial semantic/token overlap. Use when the user wants to benchmark on QRCD, or asks about evaluating this task. Reports pRR.

qhjqhj00 8a0331a 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/qrcd-eval commit 8a0331a1ea

Frequently asked questions

npx skillmds add qhjqhj00/qrcd-eval