Video Feed Ocr

Extract visible text from local video files by sampling frames and running OCR via a local multimodal LLM. Use when requests involve reading on-screen text, burned-in subtitles, slides, code snippets, captions, signs, or UI text from mp4/mov/mkv/webm recordings, reels, screen captures, or video feeds.

ikatkov 87e31f2 2 files · 14.3 KB Updated

File contents

ikatkov/agent-skills/tree/main/skills/video-feed-ocr commit 87e31f204f

Frequently asked questions

npx skillmds@latest add ikatkov/video-feed-ocr