Lrs3 Avger Eval

This evaluation probes the ability of a generative error correction model to refine audio-visual speech recognition transcripts under varying noise conditions. It measures how effectively multimodal cues (lip video and audio) combined with N-best hypotheses can reduce transcription errors compared to baseline systems. Use when the user wants to benchmark on LRS3, or asks about evaluating this task. Reports WER.

qhjqhj00 e760986 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/lrs3-avger-eval commit e7609862e9

Frequently asked questions

npx skillmds add qhjqhj00/lrs3-avger-eval