Transcribe

This skill should be used when the user wants to "transcribe audio", "transcribe video", "speech to text", "convert audio to text", "transcription", or needs to extract text from audio/video files. Default engine is 妙记 (Volcano Lark Minutes ASR) for best speaker diarization; Gemini/OpenAI available as fallbacks. Supports files up to ~8.4 hours with speaker diarization.

xingfanxia 9380940 5 files · 56.8 KB Updated

File contents

xingfanxia/ax-skills/tree/main/transcribe commit 9380940531

Frequently asked questions

npx skillmds@latest add xingfanxia/transcribe