Evalscope Docs

USE THIS SKILL WHEN working with EvalScope (ModelScope LLM evaluation framework): running evaluations, TaskConfig, supported datasets/benchmarks, evaluation backends (Native/OpenCompass/VLMEvalKit/RAGEval), performance stress testing (perf), custom datasets, multi-modal eval, arena mode, visualization, or integrating with vLLM/Swift/SGLang. Triggers on: evalscope, EvalScope, run_task, TaskConfig, evalscope eval, evalscope perf, ModelScope eval.

wenerme 502f3c0 346 files · 2.1 MB Updated

File contents

wenerme/ai/tree/main/skills/evalscope-docs commit 502f3c0be9

Frequently asked questions

npx skillmds@latest add wenerme/evalscope-docs