# Mistral Stt

> Transcribe audio files using Mistral AI Voxtral.

- Skill: `llblab/mistral-stt` (Agent Skill, multi-file: 5 files)
- Install (CLI): `npx skillmds@latest add llblab/mistral-stt`
- Raw SKILL.md: https://api.skillmd.com/api/skills/llblab/mistral-stt/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: Coding & Dev Tools
- Author: llblab (https://skillmd.com/u/llblab)
- Updated: 2026-09-10
- Page: https://skillmd.com/skills/llblab/mistral-stt

---


# Mistral STT Skill

Standalone direct Node.js client for Mistral's Voxtral transcription API. The canonical client is `scripts/transcribe.mjs`; `scripts/transcribe.sh` is the shell entrypoint wrapper. There are no curl fallbacks or Python parser dependencies.

## Usage

```bash
MISTRAL_API_KEY=xxx ./scripts/transcribe.sh audio.ogg [language] [model] [diarize]
MISTRAL_API_KEY=xxx ./scripts/transcribe.sh --file audio.ogg --lang ru --model voxtral-mini-latest --diarize true
```

- Outputs plain transcription text, or timestamped speaker segments when diarization is enabled.
- Fails fast when the file or `MISTRAL_API_KEY` is missing.

## CLI Options

- `--file`, `-f` — audio file path.
- `--lang`, `--language`, `-l` — optional language code; omitted means provider auto-detection.
- `--model`, `-m` — Mistral transcription model; default: `voxtral-mini-latest`.
- `--diarize`, `-d` — `true` to label speaker segments; default: `false`.
- `--help`, `-h` — usage.

## Dependencies

- Node.js 18+ with built-in `fetch`, `FormData`, and `Blob`.
- Internet access.

## Notes

- Diarization preserves Mistral's `speaker_id` values without guessing or merging speakers.

