Model Tester

Run reproducible local or API-based AI model evaluations across text, architecture, vision, and llama-server speed tasks, preserving per-item responses and source-to-leaderboard traceability. Use when adding a model, rerunning benchmarks, comparing models, diagnosing suspicious scores, or publishing model-evaluation results.

yenhao-huang Updated

File contents

yenhao-huang/mcp-skills-package/tree/main/skills/custom/ai-models/model-tester commit 7f9dfe00e5

Frequently asked questions

npx skillmds@latest add yenhao-huang/model-tester