Zhuangbench Eval

Evaluates large language models' ability to perform zero-shot machine translation into completely unseen, low-resource languages (Zhuang and Kalamang) using in-context learning. It probes how effectively models can adapt to new languages without prior training data by leveraging lexical expansion and syntactic exemplar retrieval. Use when the user wants to benchmark on ZhuangBench, MTOB, or asks about evaluating this task. Reports BLEU.

qhjqhj00 6624bea 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/zhuangbench-eval commit 6624bea051

Frequently asked questions

npx skillmds add qhjqhj00/zhuangbench-eval