Plugins
1 pluginResults for “speech”
43 skillsfal-audio
Convert text to speech and speech to text using fal.ai audio models.
42.4k
vox
Runs a local voice MCP server in Rust for text-to-speech and speech-to-text, with build, test, and configuration guidance.
54 · bundle
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
2
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
More results
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
0
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
6
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
2
text-to-speech
Generate natural speech from text using ElevenLabs voice AI, supporting 70+ languages, multiple models, and various output formats.
363 · bundle
speech-is-cheap-sic-skill
Fast, accurate, and incredibly inexpensive automatic speech-to-text transcription service.
12 · bundle
daily
Reference for building real-time voice and multimodal AI applications with Pipecat, covering pipelines, speech services, LLM integration, and transports.
253
nemotron-speech
Routes NVIDIA Nemotron Speech (Riva) NIM tasks for ASR, TTS, and NMT, covering cloud-hosted inference, self-hosted Docker deployment, and custom model builds.
2.2k · bundle
daily
Reference for building real-time voice and multimodal AI agents with Pipecat, covering pipelines, speech services, LLMs, transports, and deployment.
2
speech-to-text
Transcribe audio to text using ElevenLabs Scribe v2, supporting 90+ languages, speaker diarization, and word-level timestamps.
363 · bundle
sag
ElevenLabs text-to-speech with mac-style say UX.
9
sag
ElevenLabs text-to-speech with mac-style say UX.
0
sag
ElevenLabs text-to-speech with mac-style say UX.
2 · bundle
sag
ElevenLabs text-to-speech with mac-style say UX.
0
sag
ElevenLabs text-to-speech with mac-style say UX.
0
sag
ElevenLabs text-to-speech with mac-style say UX.
228
sag
ElevenLabs text-to-speech with mac-style say UX.
0 · bundle
lvxiucai-skill
吕秀才(情景喜剧虚构)认知与表达框架(压缩蒸馏):子曰嘴炮、读书人迂阔与意外高光 触发:武林外传 等。虚构
9 · bundle
sag
ElevenLabs text-to-speech with mac-style say UX.
0
transcribe
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
0 · bundle
cw-dialogue
Write natural, subtext-rich dialogue that sounds distinct per character.
0
resume-session
セッション再開時に最新5件のhandoffを読み込み文脈を復元するスキル。「おはよう」「こんにちは」「こんばんは」「再開」「restart」または /resume-session を呼んだ時にトリガーする。new-session の対(読込側)。
0
speed
Launch RSVP speed reader for text
2
using-superpowers
Use when starting any conversation - establishes how to find and use skills, requiring skill invocation before ANY response including clarifying questions
0 · bundle
transcribe
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
65 · bundle
asr-whisper-for-video-transcription-arxiv-2212-04356v1
ASR: Whisper for Video Transcription
6
ai-podcast-creation
Create AI-powered podcasts and audio content using text-to-speech, music generation, and audio editing via the inference.sh CLI.
584
asr
Transcribe audio files to text using local speech recognition. Triggers on: "转录", "transcribe", "语音转文字", "ASR", "识别音频", "把这段音频转成文字".
0 · bundle
transcribe
Transcribe audio files to text with optional speaker diarization and known-speaker hints using OpenAI models.
23.3k · bundle
transformers-js
Run state-of-the-art machine learning models directly in JavaScript/TypeScript across browsers and server-side runtimes using Transformers.js.
10.8k · bundle