OpenAI Whisper Batch Transcription Pipeline

Processes audio files from an S3 bucket using Whisper large-v3, splitting recordings into 30-second chunks with ffmpeg before transcription. Outputs timestamped SRT and VTT subtitle files plus plain-text transcripts, then uploads artifacts back to S3. Supports language auto-detection and translation to English.

agentskillexchange Updated 28 repo stars

File contents

agentskillexchange/skills/tree/main/skills/whisper-batch-transcription-pipeline commit d4411b5cd9

Frequently asked questions

npx skillmds@latest add agentskillexchange/openai-whisper-batch-transcription-pipeline