Results for “automatic-speech-recognition”

18 skills
github
Resemble Detect
Detect AI-generated audio, images, video, and text, trace synthesis sources, apply watermarks, verify speaker identity, and analyze media intelligence using the Resemble AI platform.
36.2k · bundle
modbender
Speech Is Cheap Sic Skill
Fast, accurate, and incredibly inexpensive automatic speech-to-text transcription service.
12 · bundle
mukul975
Detecting Deepfake Audio In Vishing Attacks
Detects AI-generated deepfake audio used in voice phishing (vishing) attacks by extracting spectral features and classifying samples with machine learning models.
24.6k · bundle
majiayu000
Asr
Transcribe audio files to text using the z-ai-web-dev-sdk, with CLI and SDK examples for single files, batches, and directories.
567 · bundle
jiachen-t-wang
Asr Whisper For Video Transcription Arxiv 2212 04356v1
ASR: Whisper for Video Transcription
6
omer-metin
Voice Agents
Voice Agents
128 · bundle
openai
Transcribe
Transcribe audio files to text with optional speaker diarization and known-speaker hints using OpenAI models.
23.3k · bundle
sandeeprdy1729
Asr
Comprehensive guide to asr. Master the concepts, implementation, best practices, and real-world applications of asr in professional environments.
1
diegojcn
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
brycewang-stanford
Stop Slop
Remove AI writing patterns from prose. Use when drafting, editing, or reviewing text to eliminate predictable AI tells.
1k · bundle
bouclem
Stop Slop
Remove AI writing patterns from prose. Use when drafting, editing, or reviewing text to eliminate predictable AI tells.
7 · bundle
desesbraker
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
mocchalera
Finish Interview
Applies dialogue MA, portrait reframing, and loudness/sync QA to an interview or talking-head rough cut using canonical timeline metadata and a shared renderer.
3 · bundle
rootcastleco
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
6
welitonevoc
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
mit-network
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
2
inskillflow
Fal Audio
Text-to-speech and speech-to-text using fal.ai audio models
1
jiachen-t-wang
Glip Grounded Language Image Pre Training Arxiv 2112 03857v2
GLIP: Grounded Language-Image Pre-training
6