Results for “voice-dictation”
18 skillsDialogue Audio
Create realistic multi-speaker dialogue audio using Dia TTS via the inference.sh CLI, with control over speaker tags, emotion, pacing, and conversation structure.
584
Cw Dialogue
Write natural, subtext-rich dialogue that sounds distinct per character.
0
Voice
Use when a draft needs the writer's voice and style principles applied
1
Voice Agents
Voice Agents
128 · bundle
Seed Audio
用自然语言描述生成目标音频。把一段场景描述(人声对话、环境声、音效、背景音乐等复合音频)一次性生成成音频。当用户描述一个声音场景、要求生成/合成/制作一段音频或声音、给出形如"角色:台词"的对话脚本要转成音频、或要按参考音频的音色说话时使用。支持两种模式:纯文本描述生成(T2A)和带参考音频生成(A2A,在描述中引用参考音频指定角色音色)
9 · bundle
UX Voice Tone
Voice and Tone
18 · bundle
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
Connotation Cop
Police the project's vocabulary — bust vague terms, keep the CONTEXT.md glossary sharp, and lock in decisions worth remembering as ADRs. Use when the user debates naming, says "what should we call this", asks to pin down terminology, wants a decision recorded, or when another skill (hot-seat, whiteboard) surfaces a decision that clears the ADR bar. Just reading the glossary for vocabulary is NOT this skill — trigger only when the words or decisions are being changed.
0 · bundle
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Fal Audio
Convert text to speech and speech to text using fal.ai audio models.
42.4k
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
0
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
6
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
Asr Whisper For Video Transcription Arxiv 2212 04356v1
ASR: Whisper for Video Transcription
6
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Transcribe
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
0 · bundle