Whisper

OpenAI's general-purpose speech recognition model. Supports 99 languages, transcription, translation to English, and language identification. Six model sizes from tiny (39M params) to large (1550M params). Use for speech-to-text, podcast transcription, or multilingual audio processing. Best for robust, multilingual ASR.

synthetic-sciences d61885e 2 files · 11.7 KB Updated

File contents

synthetic-sciences/openscience/tree/main/backend/cli/skills/llm-tools/whisper commit d61885ea4d

Frequently asked questions

npx skillmds@latest add synthetic-sciences/whisper