Multiwoz Dialogue Eval

Evaluates the quality, diversity, and goal adherence of task-oriented dialogue generation models. It measures how well a model generates natural, diverse responses while correctly incorporating specified dialogue goals and slot values. Use when the user wants to benchmark on MultiWOZ, or asks about evaluating this task. Reports BLEU-4.

qhjqhj00 b6bb06e 4.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/multiwoz-dialogue-eval commit b6bb06e59e

Frequently asked questions

npx skillmds add qhjqhj00/multiwoz-dialogue-eval