Hugging Face Evaluation

Add, import, and manage evaluation results in Hugging Face model cards. Supports extracting eval tables from README content, importing benchmark scores from Artificial Analysis API, and running custom model evaluations with vLLM/lighteval/inspect-ai on HF Jobs or locally. Works with the model-index metadata format for leaderboard and Papers with Code integration. Always activate when the user mentions HF model card evaluation, benchmark scores, model-index YAML, evaluation results, leaderboard submission, Artificial Analysis benchmarks, lighteval, inspect-ai, running evals on HF Jobs, or adding eval metrics to a model card — even if they don't say "skill".

Yog-Sotho Updated

File contents

Yog-Sotho/claude-skills/tree/main/hugging-face-evaluation commit 2a7e6f35e7

Frequently asked questions

npx skillmds@latest add yog-sotho/hugging-face-evaluation