Wangchanthaiinstruct Eval

Evaluates instruction-following capabilities of LLMs in Thai across culture-aware, domain-specific (Medical, Law, Finance, Retail), and multitask settings. Probes factual accuracy, reasoning quality, and fluency in both zero-shot and fine-tuned regimes. Use when the user wants to benchmark on WangchanThaiInstruct, Thai LLM Leaderboard, Thai MT-Bench, or asks about evaluating this task. Reports Accuracy.

qhjqhj00 8764284 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/wangchanthaiinstruct-eval commit 8764284030

Frequently asked questions

npx skillmds add qhjqhj00/wangchanthaiinstruct-eval