Speech To Text

Transcribe audio to text using Sarvam AI's Saaras model. Handles speech recognition, transcription, and voice interfaces for 23 Indian languages. Supports 5 output modes, auto language detection, WebSocket streaming, and batch diarization. Use when converting speech to text or building voice-enabled apps. Use when this capability is needed.

tomevault-io Updated

File contents

tomevault-io/skills-registry/tree/main/sarvamai--skills--speech-to-text commit 7c1692f5c3

Frequently asked questions

npx skillmds@latest add tomevault-io/speech-to-text