Flashinfer

FlashInfer — High-performance kernel library for LLM inference with optimized attention, paged KV-cache, FP8/FP4 quantization

wenyi-li Updated

File contents

wenyi-li/awesome-agent-kernel-skills/tree/main/agent-compute-flashinfer commit 7859772f58

Frequently asked questions

npx skillmds@latest add wenyi-li/flashinfer