Cultural Nuance Mt Eval

This benchmark evaluates how well multilingual LLMs preserve cultural nuance, idioms, puns, and culturally embedded concepts during machine translation. It probes the persistent gap between grammatical accuracy and cultural resonance by measuring translation quality across different figurative and non-figurative segment categories. Use when the user wants to benchmark on Cultural Nuance MT Benchmark, or asks about evaluating this task. Reports overall quality.

qhjqhj00 2384d29 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cultural-nuance-mt-eval commit 2384d29a6a

Frequently asked questions

npx skillmds add qhjqhj00/cultural-nuance-mt-eval