Iblai API Agent Eval

Measure and improve agent quality via the platform API — evaluation datasets, dataset items (JSON, CSV upload, or from chat traces), experiment runs, LLM-as-Judge and human-annotation scoring, score configs, and CSV export. Use to test an agent against a dataset and grade the results.

iblai 05cefe5 2 files · 13.2 KB Updated

File contents

iblai/vibe/tree/main/skills/agents/iblai-api-agent-eval commit 05cefe55d5

Frequently asked questions

npx skillmds@latest add iblai/iblai-api-agent-eval