Results for “automatic-speech-recognition”
18 skillsResemble Detect
Detect AI-generated audio, images, video, and text, trace synthesis sources, apply watermarks, verify speaker identity, and analyze media intelligence using the Resemble AI platform.
36.2k · bundle
Speech Is Cheap Sic Skill
Fast, accurate, and incredibly inexpensive automatic speech-to-text transcription service.
12 · bundle
Detecting Deepfake Audio In Vishing Attacks
Detects AI-generated deepfake audio used in voice phishing (vishing) attacks by extracting spectral features and classifying samples with machine learning models.
24.6k · bundle
Asr
Transcribe audio files to text using the z-ai-web-dev-sdk, with CLI and SDK examples for single files, batches, and directories.
567 · bundle
Asr Whisper For Video Transcription Arxiv 2212 04356v1
ASR: Whisper for Video Transcription
6
Voice Agents
Voice Agents
128 · bundle
Transcribe
Transcribe audio files to text with optional speaker diarization and known-speaker hints using OpenAI models.
23.3k · bundle
Asr
Comprehensive guide to asr. Master the concepts, implementation, best practices, and real-world applications of asr in professional environments.
1
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Stop Slop
Remove AI writing patterns from prose. Use when drafting, editing, or reviewing text to eliminate predictable AI tells.
1k · bundle
Stop Slop
Remove AI writing patterns from prose. Use when drafting, editing, or reviewing text to eliminate predictable AI tells.
7 · bundle
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
Finish Interview
Applies dialogue MA, portrait reframing, and loudness/sync QA to an interview or talking-head rough cut using canonical timeline metadata and a shared renderer.
3 · bundle
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
6
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
Glip Grounded Language Image Pre Training Arxiv 2112 03857v2
GLIP: Grounded Language-Image Pre-training
6