Iblai API Agent Eval

Measure and improve agent quality via the platform API — evaluation datasets, dataset items (JSON, CSV upload, or from chat traces), experiment runs, LLM-as-Judge and human-annotation scoring, score configs, and CSV export. Use to test an agent against a dataset and grade the results.

iblai Updated

File contents

iblai/api/tree/main/skills/iblai-api-agent-eval commit 5a2bc05657

Frequently asked questions

npx skillmds@latest add iblai/iblai-api-agent-eval-2