Evaluate Multimodal

Evaluate multimodal AI agents that process images, audio, PDFs, or other files. Sets up evaluations using LangWatch's LLM-as-judge with image inputs, Scenario's multimodal testing, and document parsing evaluation patterns. Use when your agent handles non-text inputs.

langwatch cdaeea3 3.9 KB Updated

File contents

langwatch/langwatch/tree/main/skills/_compiled/native/evaluate-multimodal commit cdaeea3664

Frequently asked questions

npx skillmds@latest add langwatch/evaluate-multimodal