Dolphin Arabic Nlg Eval

Evaluates the natural language generation capabilities of models across 13 diverse Arabic tasks, including machine translation, summarization, question generation, and dialectal normalization. It probes how well models handle linguistic variability across Classical Arabic, Modern Standard Arabic, dialects, and Arabizi, as well as cross-lingual and code-switched scenarios. Use when the user wants to benchmark on Dolphin, or asks about evaluating this task. Reports BLEU.

qhjqhj00 83d1ce2 4.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/dolphin-arabic-nlg-eval commit 83d1ce25a2

Frequently asked questions

npx skillmds add qhjqhj00/dolphin-arabic-nlg-eval