Bengalimoralbench Eval

Evaluates large language models' ability to perform moral reasoning and align with human ethical judgments within Bengali language and South Asian socio-cultural contexts. It probes cultural grounding, commonsense reasoning, and fairness across five everyday moral domains using native-speaker consensus annotations. Use when the user wants to benchmark on BengaliMoralBench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 bc4c0e8 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/bengalimoralbench-eval commit bc4c0e8944

Frequently asked questions

npx skillmds add qhjqhj00/bengalimoralbench-eval