Search Benchmark

Generate a search quality benchmark for the AI Registry. Generates ground truth from the registry's assets, runs 100+ queries against the semantic search API, evaluates results using NDCG@10/MRR/Recall, and produces a markdown report. Use when you want to measure search quality after changes to the scoring algorithm, embedding model, or indexed content.

agentic-community Updated

File contents

agentic-community/mcp-gateway-registry/tree/main/.claude/skills/search-benchmark commit 1df7c37e81

Frequently asked questions

npx skillmds@latest add agentic-community/search-benchmark