Whisper.cpp Local Transcription Engine

Runs OpenAI Whisper models locally via whisper.cpp with GGML quantized weights for CPU-efficient transcription. Supports beam search decoding, VAD-based segmentation, and SRT/VTT subtitle output formats.

agentskillexchange Updated 28 repo stars

File contents

Whisper.cpp Local Transcription Engine

Runs OpenAI Whisper models locally via whisper.cpp with GGML quantized weights for CPU-efficient transcription. Supports beam search decoding, VAD-based segmentation, and SRT/VTT subtitle output formats.

Installation

Use the upstream whisper.cpp build path for local transcription:

git clone https://github.com/ggml-org/whisper.cpp.git
cd whisper.cpp
sh ./models/download-ggml-model.sh base.en
cmake -B build
cmake --build build -j --config Release
./build/bin/whisper-cli -f samples/jfk.wav

For a quick demo, upstream also supports:

make base.en

The whisper-cli example currently expects 16-bit WAV input. Convert other audio formats before transcription when needed:

ffmpeg -i input.mp3 -ar 16000 -ac 1 -c:a pcm_s16le output.wav
./build/bin/whisper-cli -f output.wav

Run ./build/bin/whisper-cli -h for detailed CLI options.

Source

agentskillexchange/skills/tree/main/skills/whisper-cpp-local-transcription-engine commit e411dec6f9

Frequently asked questions

npx skillmds@latest add agentskillexchange/whisper-cpp-local-transcription-engine