Magpie Kernel Evaluator

Benchmarks LLM inference and drives GPU kernel optimization with Magpie. Use when the user wants to benchmark vLLM, SGLang, or Atom; capture torch traces; post-process inference traces with TraceLens into prefill/decode and roofline reports; identify top bottleneck kernels or map profiler names to source; analyze or compare HIP, CUDA, PyTorch, or Triton kernels; validate and rank optimized variants; run local, container, or Ray workloads; or mentions Magpie, TraceLens, gap analysis, TTFT, TPOT, kernel evaluation, or AMD GPU optimization.

amd bbcc2ee 10 files · 33.3 KB Updated

File contents

amd/skills/tree/main/skills/magpie-kernel-evaluator commit bbcc2ee7ca

Frequently asked questions

npx skillmds@latest add amd/magpie-kernel-evaluator