Edgespeak Align

Force-align audio/video against a known transcript on-device via EdgeSpeak to produce word-level timestamps (start, end, score) for karaoke captions, word-accurate SRT, dubbing, and clip extraction. Use when the user already has the transcript/script/lyrics and wants to know exactly when each word is spoken.

lattifai 8030b08 9.6 KB Updated

File contents

lattifai/edgespeak-skills/tree/main/skills/edgespeak-align commit 8030b0804c

Frequently asked questions

npx skillmds@latest add lattifai/edgespeak-align