Vision Augment

视觉增强 MCP —— 为无视觉能力的 LLM(如 DeepSeek、GLM)提供看图、OCR 与文档解析。当用户需要分析图片/截图内容、识别图片中的文字、解析 docx/pdf/pptx/xlsx/html 文档时使用。工具:mcp_vision_augment_vision(task_type=reasoning|ocr|document)、mcp_vision_augment_health(环境探测)、mcp_vision_augment_clear_cache。

caomeiyouren 1fba831 11 files · 577.5 KB Updated

File contents

caomeiyouren/vision-augment/tree/main/ commit 1fba8311c5

Frequently asked questions

npx skillmds@latest add caomeiyouren/vision-augment