Hkmmlu Eval

Evaluates large language models' multilingual comprehension of Hong Kong-specific knowledge, Cantonese linguistic capabilities, and reasoning across STEM, social sciences, and humanities in both Traditional and Simplified Chinese. Use when the user wants to benchmark on HKMMLU, or asks about evaluating this task. Reports accuracy.

qhjqhj00 00eecd7 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/hkmmlu-eval commit 00eecd79d9

Frequently asked questions

npx skillmds add qhjqhj00/hkmmlu-eval