Speakstream Streaming Tts Eval

Evaluates the quality and latency of streaming text-to-speech models that generate audio incrementally from interleaved text and speech inputs. It probes the model's ability to maintain speech accuracy and naturalness while minimizing first-token latency under streaming constraints. Use when the user wants to benchmark on LJSpeech, LibriSpeech, or asks about evaluating this task. Reports Word Error Rate (WER).

qhjqhj00 08290a2 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/speakstream-streaming-tts-eval commit 08290a2837

Frequently asked questions

npx skillmds add qhjqhj00/speakstream-streaming-tts-eval