Results for “vocal-isolation”

49 skills
More results
elevenlabs
voice-isolator
Remove background noise and isolate vocals or speech from audio files using the ElevenLabs Voice Isolator API.
363 · bundle
smith6jt-cop
batch-isolation
Batch Signal Isolation with Recipe-Driven Processing
3
ichichuang
songsee
Generate spectrograms and audio feature visualizations (mel, chroma, MFCC, tempogram, etc.) from audio files via CLI. Useful for audio analysis, music production debugging, and visual documentation.
0 · bundle
intense-visions
ux-voice-tone
Voice and Tone
18 · bundle
qhjqhj00
sdr
Quantifies audio source separation quality by computing the signal-to-distortion ratio (SDR) between ground-truth and estimated stems, with per-stem and record-level averaging.
3
loopyluci
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
1
lord1egypt
songsee
Generates spectrograms and multi-panel audio feature visualizations (mel, chroma, MFCC) from audio files via a Go CLI.
2
oyi77
voice-ai
Generates speech, transcribes audio, clones voices, and builds real-time voice agents using ElevenLabs, OpenAI TTS, Whisper, and Vapi.
10
q2805187159
songsee
Generate spectrograms and audio feature visualizations (mel, chroma, MFCC, tempogram, etc.) from audio files via CLI. Useful for audio analysis, music production debugging, and visual documentation.
3
diegojcn
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
1
aniruddhaadak80
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
0
elevenlabs
voice-changer
Transform the voice in an audio recording into a different target voice while preserving emotion, timing, and delivery using the ElevenLabs Voice Changer API.
363 · bundle
construct-ai-primary
voice-maestro
Use when voice AI strategy, conversational AI architecture, voice technology innovation, or voice platform leadership is needed. This agent specializes in voice AI leadership within the VoiceForge AI ecosystem.
0
bog5d
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
0
ziri22
agent-songsee-v2
Expert en analyse audio avancé (spectrograms, mel, chroma, MFCC, feature extraction, CLI)
6
ekatasingh1107
brand-voice
Extract and codify brand voice from existing content for consistent messaging
2 · bundle
fukukei23
analyze-song
楽曲(YouTube/MP3)を音源から定量分析し、BPM/キー/コード進行/メロディ輪郭/音域/phrase_repetitionを抽出してfeatures.json+五線譜PNG/PDF+report.mdを出力するスキル。reverse-engineer-song(Gemini定性)とは完全独立・数値定量分析専用。ユーザーが「楽曲分析して」「曲を定量分析」「この曲のBPM/コード抽出」「名曲っぽさ分析」「analyze-song」と言った時、または /analyze-song を呼んだ時にトリガー。
0 · bundle
netanel-abergel
whatsapp-voice
Transcribe WhatsApp voice messages using local Whisper CLI. Use when: owner or contact sends an audio/ogg voice message. Combines Whisper transcription + CRM update + task creation. Works offline for short clips, uses OpenAI API for long clips. Hebrew and English supported.
6
doriangallo
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
ziri22
agent-voice-search
Expert en optimisation recherche vocale (structured data, featured snippets, queries conversationnelles)
6
samyakjhaveri
overnight-eval
Launches long-running evaluation batches in isolated tmux sessions with pre-flight verification, monitoring, and post-flight analysis for unattended runs.
0
loopyluci
buddy-sings
**Trigger**: Use when the user wants their AI pet or buddy to sing a personalized song. Creates a unique vocal identity, gathers conversation context, and generates custom music.
1
inference-sh
ai-voice-cloning
Generate natural AI voices, text-to-speech, and voice synthesis using the inference.sh CLI with models like Inworld TTS, ElevenLabs, and Kokoro TTS for voiceovers, audiobooks, podcasts, and more.
584
qhjqhj00
caa-eval
Benchmarks large audio-language models against adversarial audio attacks using the CAA dataset, computing WER, ROUGE-L, cosine similarity, and coherence scores to assess robustness in conversational settings.
3
rootcastleco
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
6
lucian55
beidao-skill
北岛(诗人)认知与表达框架(压缩蒸馏):冷意象、否定句诗学、流亡与记忆 触发:回答、朦胧诗 等。不伪造诗句
9 · bundle
promisingcoder
songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
0
owl-listener
von-restorff-effect
Apply the Von Restorff Effect to make the most important element visually distinct from its surroundings, improving attention and recall.
1.7k
smith6jt-cop
multi-project-batch-isolation
Multi-project signal isolation with cascading recipe resolution
3
iamanacarolinarezende
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
0
promisingcoder
sonoscli
Control Sonos speakers (discover/status/play/volume/group).
0
diegojcn
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
jackychenlu
songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
0
orchestra-research
whisper
Transcribe and translate speech across 99 languages using OpenAI's Whisper model, with support for multiple model sizes, batch processing, and subtitle generation.
10.4k · bundle
lucian55
pujian-skill
朴树(音乐人)认知与表达框架(压缩蒸馏):内向真诚、少话多留白、生命与创作一体… 触发:平凡之路 等。不消费抑郁;非医疗建议
9 · bundle