Klue Dst Eval

Evaluates a model's ability to track and update dialogue state across turns in Korean conversations, testing multi-turn reasoning and slot filling. Use when the user wants to benchmark on KLUE-DST, or asks about evaluating this task. Reports Joint Accuracy.

qhjqhj00 f0fec24 1.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/klue-dst-eval commit f0fec24a60

Frequently asked questions

npx skillmds add qhjqhj00/klue-dst-eval