Benchmark Research Skill

Claude Code style skill for benchmark research. Use when the user asks to analyze one paper's benchmarks, datasets, metrics, experiment tables/figures, baselines, or related work; or asks to survey a direction and find usable benchmarks for evaluation. Helper scripts fetch paper context/assets and Claude Code performs the semantic extraction.

EternalWavee Updated

File contents

EternalWavee/benchmark-research-skill/tree/main/ commit 5fd9d9c96b

Frequently asked questions

npx skillmds@latest add eternalwavee/benchmark-research-skill