Automatic Speech Recognition Asr

Transcribe audio segments to text using Whisper models. Use larger models (small, base, medium, large-v3) for better accuracy, or faster-whisper for optimized performance. Always align transcription timestamps with diarization segments for accurate speaker-labeled subtitles. Use when this capability is needed.

tomevault-io Updated

File contents

tomevault-io/skills-registry/tree/main/benchflow-ai--skillsbench--automatic-speech-recognition commit ad007f0293

Frequently asked questions

npx skillmds@latest add tomevault-io/automatic-speech-recognition-asr