Plugins

1 plugin

Results for “speech”

43 skills
More results
iamanacarolinarezende
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
0
doriangallo
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
rootcastleco
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
6
mit-network
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
2
elevenlabs
text-to-speech
Generate natural speech from text using ElevenLabs voice AI, supporting 70+ languages, multiple models, and various output formats.
363 · bundle
modbender
speech-is-cheap-sic-skill
Fast, accurate, and incredibly inexpensive automatic speech-to-text transcription service.
12 · bundle
lingxling
daily
Reference for building real-time voice and multimodal AI applications with Pipecat, covering pipelines, speech services, LLM integration, and transports.
253
nvidia
nemotron-speech
Routes NVIDIA Nemotron Speech (Riva) NIM tasks for ASR, TTS, and NMT, covering cloud-hosted inference, self-hosted Docker deployment, and custom model builds.
2.2k · bundle
nimoqup046-collab
daily
Reference for building real-time voice and multimodal AI agents with Pipecat, covering pipelines, speech services, LLMs, transports, and deployment.
2
elevenlabs
speech-to-text
Transcribe audio to text using ElevenLabs Scribe v2, supporting 90+ languages, speaker diarization, and word-level timestamps.
363 · bundle
x402agent
sag
ElevenLabs text-to-speech with mac-style say UX.
9
solizardking
sag
ElevenLabs text-to-speech with mac-style say UX.
0
danstrem2
sag
ElevenLabs text-to-speech with mac-style say UX.
2 · bundle
jackychenlu
sag
ElevenLabs text-to-speech with mac-style say UX.
0
promisingcoder
sag
ElevenLabs text-to-speech with mac-style say UX.
0
infometa
sag
ElevenLabs text-to-speech with mac-style say UX.
228
om-scogo
sag
ElevenLabs text-to-speech with mac-style say UX.
0 · bundle
lucian55
lvxiucai-skill
吕秀才(情景喜剧虚构)认知与表达框架(压缩蒸馏):子曰嘴炮、读书人迂阔与意外高光 触发:武林外传 等。虚构
9 · bundle
jrennie99-glitch
sag
ElevenLabs text-to-speech with mac-style say UX.
0
metinduraktr-44
transcribe
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
0 · bundle
matrixx0070
cw-dialogue
Write natural, subtext-rich dialogue that sounds distinct per character.
0
fukukei23
resume-session
セッション再開時に最新5件のhandoffを読み込み文脈を復元するスキル。「おはよう」「こんにちは」「こんばんは」「再開」「restart」または /resume-session を呼んだ時にトリガーする。new-session の対(読込側)。
0
mmehdi0606
speed
Launch RSVP speed reader for text
2
neekware
using-superpowers
Use when starting any conversation - establishes how to find and use skills, requiring skill invocation before ANY response including clarifying questions
0 · bundle
seaworld008
transcribe
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
65 · bundle
jiachen-t-wang
asr-whisper-for-video-transcription-arxiv-2212-04356v1
ASR: Whisper for Video Transcription
6
inference-sh
ai-podcast-creation
Create AI-powered podcasts and audio content using text-to-speech, music generation, and audio editing via the inference.sh CLI.
584
om-scogo
asr
Transcribe audio files to text using local speech recognition. Triggers on: "转录", "transcribe", "语音转文字", "ASR", "识别音频", "把这段音频转成文字".
0 · bundle
openai
transcribe
Transcribe audio files to text with optional speaker diarization and known-speaker hints using OpenAI models.
23.3k · bundle
huggingface
transformers-js
Run state-of-the-art machine learning models directly in JavaScript/TypeScript across browsers and server-side runtimes using Transformers.js.
10.8k · bundle