Kernel Microbenchmark

Build, debug, and interpret vLLM GPU kernel microbenchmarks for CUDA, Triton, and CuteDSL, including CUPTI timing, correctness checks, generated-code inspection, multi-GPU measurements, and SOL sanity checks.

vllm-project f2dd5be 4 files · 17.8 KB Updated

File contents

vllm-project/vllm commit f2dd5be929

Frequently asked questions

npx skillmds add vllm-project/kernel-microbenchmark