Eval Dataset Research Prompts 30

Use when running the legal research benchmark that tests statute lookup, case law retrieval, multi-jurisdictional comparison, and regulator guidance queries across DIFC/ADGM/UK/FR/MENA. Hallucination rate is the primary metric — any fabricated citation is an automatic fail. Contains 30 prompts with a mandatory <1% hallucination target.

sboghossian Updated

File contents

sboghossian/mini-claude-for-legal/tree/main/skills/eval/eval-dataset-research-prompts-30 commit 92a8410365

Frequently asked questions

npx skillmds@latest add sboghossian-mini-claude-for-legal/eval-dataset-research-prompts-30