Masalbench Benchmark Contextual Cross Cultural

Build cross-cultural figurative language benchmarks and evaluation pipelines for LLMs. Applies the MasalBench methodology to test whether models truly understand proverbs, idioms, and culturally-embedded expressions -- not just pattern-match surface text. Trigger phrases: 'evaluate LLM cultural understanding', 'benchmark figurative language', 'test proverb comprehension', 'cross-cultural NLP evaluation', 'build idiom benchmark', 'low-resource language evaluation'

ndpvt-web Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/masalbench-benchmark-contextual-cross-cultural commit 1dc46615fd

Frequently asked questions

npx skillmds@latest add ndpvt-web/masalbench-benchmark-contextual-cross-cultural