Packs

2 packs

Results for “voice-ai”

34 skills
More results
nimoqup046-collab
daily
Reference for building real-time voice and multimodal AI agents with Pipecat, covering pipelines, speech services, LLMs, transports, and deployment.
2
inference-sh
talking-head-production
Create talking head videos with AI avatars, lipsync, and voiceover using the inference.sh CLI.
584
heath-gtm
humanize-draft
Strip AI-generated tells from a draft and rewrite it in your own voice. Fire on "humanize this", "de-AI this", "clean up this draft", "rewrite in my voice", "make this sound like me", "kill the AI tells", or "polish this draft". Applies a 30-pattern catalog of AI-writing tells, then re-injects a voice you define so the result reads like a person, not a politer machine. Pattern catalog based on Wikipedia's "Signs of AI writing." Forked from blader/humanizer (MIT).
0 · bundle
elevenlabs
text-to-speech
Generate natural speech from text using ElevenLabs voice AI, supporting 70+ languages, multiple models, and various output formats.
363 · bundle
lingxling
daily
Reference for building real-time voice and multimodal AI applications with Pipecat, covering pipelines, speech services, LLM integration, and transports.
253
diegojcn
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
desesbraker
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
2
rootcastleco
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
6
welitonevoc
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
doriangallo
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
minimax-ai
buddy-sings
Gives your Claude Code pet a unique singing voice based on its personality, generating music with MiniMax's API and playing it back.
12.9k
iamanacarolinarezende
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
0
mit-network
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
2
omer-metin
voice-agents
Voice Agents
128 · bundle
microsoft
podcast-generation
Generate AI-powered podcast-style audio narratives from text using Azure OpenAI's GPT Realtime Mini model via WebSocket, with full-stack implementation from React frontend to Python FastAPI backend.
2.7k · bundle
mukul975
detecting-deepfake-audio-in-vishing-attacks
Detects AI-generated deepfake audio used in voice phishing (vishing) attacks by extracting spectral features and classifying samples with machine learning models.
24.6k · bundle
inference-sh
ai-podcast
Generate multi-person talking head podcast videos from scratch using AI — character creation, TTS, avatar animation, and video stitching.
584
antigravity
fal-audio
Convert text to speech and speech to text using fal.ai audio models.
42.4k
inskillflow
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
intelli-verse-x
ivx-qv-audio-video
Work with AudioSource, VideoPlayer, audio quiz playback, AI voice, and media streaming in QuizVerse.
0 · bundle
yanacuti1121
brand-voice
Build a source-derived writing style profile from real posts, essays, launch notes, docs, or site copy, then reuse that profile across content, outreach, and social workflows. Use when the user wants voice consistency without generic AI writing tropes.
2
tangchunwu
brand-voice
Build a source-derived writing style profile from real posts, essays, launch notes, docs, or site copy, then reuse that profile across content, outreach, and social workflows. Use when the user wants voice consistency without generic AI writing tropes.
1 · bundle
livelybug
brand-voice
Build a source-derived writing style profile from real posts, essays, launch notes, docs, or site copy, then reuse that profile across content, outreach, and social workflows. Use when the user wants voice consistency without generic AI writing tropes.
0 · bundle
anantha-236
brand-voice
Build a source-derived writing style profile from real posts, essays, launch notes, docs, or site copy, then reuse that profile across content, outreach, and social workflows. Use when the user wants voice consistency without generic AI writing tropes.
1 · bundle
brycewang-stanford
g6
VS-Enhanced Academic Style Humanizer - Transforms writing patterns to achieve authentic scholarly voice Applies transformations from G5 analysis to create natural academic prose Use when: improving AI-assisted writing quality, preparing manuscripts, enhancing scholarly voice Triggers: humanize, transform, make natural, improve writing quality, improve style
1k
onourimpram
anti-ai-trace-revision
Use when a draft in Turkish, English, or both reads as AI-generated and needs revision that removes machine writing patterns and translation calques while preserving meaning, citations, statistics, and the author's voice.
2
doany-ai
lipsync
Lip-sync a face to a specific audio track on RunComfy via the `runcomfy` CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar from a portrait + audio), Sync Labs sync v2 / Pro (state-of-the-art mouth sync onto a video), Kling lipsync (audio-to- video and text-to-video with synced speech), and Creatify lipsync. The skill picks the right endpoint for the user's actual intent — portrait still + audio (avatar-style), source video + audio (mouth- swap on existing footage), or generate-and-sync from a script. Triggers on "lip sync", "lipsync", "make this video speak", "match audio to mouth", "dub video", "sync lips to voice", "Sync Labs", "voiceover sync", or any explicit ask to drive a face's mouth from an audio track.
5
kk20300113-png
open-gstack-browser
Launch GStack Browser — AI-controlled Chromium with the sidebar extension baked in. Opens a visible browser window where you can watch every action in real time. The sidebar shows a live activity feed and chat. Anti-bot stealth built in. Use when asked to "open gstack browser", "launch browser", "connect chrome", "open chrome", "real browser", "launch chrome", "side panel", or "control my browser". Voice triggers (speech-to-text aliases): "show me the browser".
0 · bundle