Gemini Tts

Generate read-aloud audio (text-to-speech) using Google Gemini TTS (gemini-3.1-flash-tts-preview). Automatically detects the mode from the input: single-speaker narration for plain text, and multi-speaker dialogue when the input has two "Name:" speaker labels. Supports 30 prebuilt voices, natural-language style control and audio tags, text or file input, and WAV output. Works with both the Gemini Developer API and Vertex AI.

danishi e1d4c10 4 files · 31.0 KB Updated

File contents

danishi/claude-code-config/tree/main/skills/gemini-tts commit e1d4c10f2c

Frequently asked questions

npx skillmds@latest add danishi/gemini-tts