Audio Whisper Transcription

Transcribes audio or video to text with word-level timestamps via faster-whisper (default large-v3-turbo) or local openai-whisper (word_timestamps=True). Use when the user needs captions, filler-word cut lists, or searchable transcripts with start/end times. Not for speaker diarization (pyannote/WhisperX), denoising, or hosts that ban ffmpeg. Never invent package versions such as Whisper 2024.5.

gabrielmoreira Updated 17 repo stars

File contents

gabrielmoreira/agent-skills-mirror/tree/main/mirrors/repos/Kayforkind@skill-slice/skills/audio-whisper-transcription commit cdef3078b6

Frequently asked questions

npx skillmds@latest add gabrielmoreira/audio-whisper-transcription