Chinese LLM Bench Eval

Evaluates Chinese large language models' world knowledge, academic understanding, and multi-dimensional alignment after pretraining or instruction fine-tuning. It probes the model's ability to follow instructions, reason across domains, and maintain safety and helpfulness standards in Chinese. Use when the user wants to benchmark on C-Eval, CMMLU, Alignbench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 6dde522 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/chinese-llm-bench-eval commit 6dde522ea5

Frequently asked questions

npx skillmds add qhjqhj00/chinese-llm-bench-eval