Medical Dialogue Gen Eval

This evaluation probes a model's ability to generate clinically accurate, fluent, and context-aware physician responses in multi-turn medical dialogues. It assesses both automatic language quality and semantic relevance, alongside human-rated fluency, knowledge correctness, and overall satisfaction. Use when the user wants to benchmark on KaMed, MedDialog, MedDG, or asks about evaluating this task. Reports BLEU-2.

qhjqhj00 ba92bf3 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medical-dialogue-gen-eval commit ba92bf3e8c

Frequently asked questions

npx skillmds add qhjqhj00/medical-dialogue-gen-eval