Groq Stt

Transcribe audio files using Groq API (Whisper).

llblab Updated

File contents

Groq STT Skill

Standalone direct Node.js client for Groq's Whisper transcription API. The canonical client is scripts/transcribe.mjs; scripts/transcribe.sh is the shell entrypoint wrapper. There are no curl fallbacks or JSON parser dependencies.

Usage

GROQ_API_KEY=xxx ./scripts/transcribe.sh audio.ogg [language] [model] [diarize]
GROQ_API_KEY=xxx ./scripts/transcribe.sh --file audio.ogg --lang ru --model whisper-large-v3-turbo --diarize true
  • Outputs plain transcription text, or timestamped speaker segments when diarization is enabled.
  • Fails fast when the file or GROQ_API_KEY is missing.

CLI Options

  • --file, -f — audio file path.
  • --lang, --language, -l — optional language code; omitted means provider auto-detection.
  • --model, -m — Groq transcription model; default: whisper-large-v3-turbo.
  • --diarize, -dtrue to label speaker segments; default: false.
  • --help, -h — usage.

Dependencies

  • Node.js 18+ with built-in fetch, FormData, and Blob.
  • Internet access.

Notes

  • Diarization preserves Groq's speaker labels without guessing or merging speakers.

llblab/skills/tree/main/groq-stt commit 7a8ab68a06

Frequently asked questions

npx skillmds@latest add llblab/groq-stt