Results for “vocal-isolation”
49 skillsMore results
voice-isolator
Remove background noise and isolate vocals or speech from audio files using the ElevenLabs Voice Isolator API.
363 · bundle
batch-isolation
Batch Signal Isolation with Recipe-Driven Processing
3
songsee
Generate spectrograms and audio feature visualizations (mel, chroma, MFCC, tempogram, etc.) from audio files via CLI. Useful for audio analysis, music production debugging, and visual documentation.
0 · bundle
ux-voice-tone
Voice and Tone
18 · bundle
sdr
Quantifies audio source separation quality by computing the signal-to-distortion ratio (SDR) between ground-truth and estimated stems, with per-stem and record-level averaging.
3
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
1
songsee
Generates spectrograms and multi-panel audio feature visualizations (mel, chroma, MFCC) from audio files via a Go CLI.
2
voice-ai
Generates speech, transcribes audio, clones voices, and builds real-time voice agents using ElevenLabs, OpenAI TTS, Whisper, and Vapi.
10
songsee
Generate spectrograms and audio feature visualizations (mel, chroma, MFCC, tempogram, etc.) from audio files via CLI. Useful for audio analysis, music production debugging, and visual documentation.
3
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
1
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
0
voice-changer
Transform the voice in an audio recording into a different target voice while preserving emotion, timing, and delivery using the ElevenLabs Voice Changer API.
363 · bundle
voice-maestro
Use when voice AI strategy, conversational AI architecture, voice technology innovation, or voice platform leadership is needed. This agent specializes in voice AI leadership within the VoiceForge AI ecosystem.
0
songsee
Audio spectrograms/features (mel, chroma, MFCC) via CLI.
0
agent-songsee-v2
Expert en analyse audio avancé (spectrograms, mel, chroma, MFCC, feature extraction, CLI)
6
brand-voice
Extract and codify brand voice from existing content for consistent messaging
2 · bundle
analyze-song
楽曲(YouTube/MP3)を音源から定量分析し、BPM/キー/コード進行/メロディ輪郭/音域/phrase_repetitionを抽出してfeatures.json+五線譜PNG/PDF+report.mdを出力するスキル。reverse-engineer-song(Gemini定性)とは完全独立・数値定量分析専用。ユーザーが「楽曲分析して」「曲を定量分析」「この曲のBPM/コード抽出」「名曲っぽさ分析」「analyze-song」と言った時、または /analyze-song を呼んだ時にトリガー。
0 · bundle
whatsapp-voice
Transcribe WhatsApp voice messages using local Whisper CLI. Use when: owner or contact sends an audio/ogg voice message. Combines Whisper transcription + CRM update + task creation. Works offline for short clips, uses OpenAI API for long clips. Hebrew and English supported.
6
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
agent-voice-search
Expert en optimisation recherche vocale (structured data, featured snippets, queries conversationnelles)
6
overnight-eval
Launches long-running evaluation batches in isolated tmux sessions with pre-flight verification, monitoring, and post-flight analysis for unattended runs.
0
buddy-sings
**Trigger**: Use when the user wants their AI pet or buddy to sing a personalized song. Creates a unique vocal identity, gathers conversation context, and generates custom music.
1
ai-voice-cloning
Generate natural AI voices, text-to-speech, and voice synthesis using the inference.sh CLI with models like Inworld TTS, ElevenLabs, and Kokoro TTS for voiceovers, audiobooks, podcasts, and more.
584
caa-eval
Benchmarks large audio-language models against adversarial audio attacks using the CAA dataset, computing WER, ROUGE-L, cosine similarity, and coherence scores to assess robustness in conversational settings.
3
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
6
beidao-skill
北岛(诗人)认知与表达框架(压缩蒸馏):冷意象、否定句诗学、流亡与记忆 触发:回答、朦胧诗 等。不伪造诗句
9 · bundle
songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
0
von-restorff-effect
Apply the Von Restorff Effect to make the most important element visually distinct from its surroundings, improving attention and recall.
1.7k
multi-project-batch-isolation
Multi-project signal isolation with cascading recipe resolution
3
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
0
sonoscli
Control Sonos speakers (discover/status/play/volume/group).
0
fal-audio
Text-to-speech and speech-to-text using fal.ai audio models
1
songsee
Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.
0
whisper
Transcribe and translate speech across 99 languages using OpenAI's Whisper model, with support for multiple model sizes, batch processing, and subtitle generation.
10.4k · bundle
pujian-skill
朴树(音乐人)认知与表达框架(压缩蒸馏):内向真诚、少话多留白、生命与创作一体… 触发:平凡之路 等。不消费抑郁;非医疗建议
9 · bundle