Visual Rpa

Visual RPA desktop automation skill. Use when user asks to operate desktop apps, click icons, open applications, type text in input fields, click buttons, scroll pages, send messages via WeChat or other apps. Uses screen capture and Qwen vision model for pure visual positioning without DOM or accessibility APIs.

modbender fc5908f 2 files · 29.3 KB Updated 12 repo stars

File contents

modbender/skill-library-mcp/tree/main/data/visual-rpa-skill commit fc5908f593

Frequently asked questions

npx skillmds add modbender/visual-rpa