Towards Automated Kernel Generation

Automate GPU kernel generation and optimization using LLM-driven agentic workflows with profiling feedback loops. Use when user asks to 'write a CUDA kernel', 'optimize a Triton kernel', 'generate a GPU kernel for this operation', 'speed up this PyTorch operator with a custom kernel', 'convert this PyTorch op to Triton', or 'profile and optimize my kernel'.

ndpvt-web 7310fb6 14.4 KB Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/towards-automated-kernel-generation commit 7310fb6a46

Frequently asked questions

npx skillmds@latest add ndpvt-web/towards-automated-kernel-generation