Profinfer Ebpf Based Fine Grained Inference

Profile and diagnose LLM inference engines (llama.cpp and similar GGML-based runtimes) using eBPF uprobes for non-intrusive, operator-level performance analysis. Trigger phrases: 'profile llama.cpp inference', 'eBPF LLM profiling', 'diagnose inference bottleneck', 'trace GGML operators', 'is my inference memory-bound or compute-bound', 'profile MoE routing overhead'

ndpvt-web Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/profinfer-ebpf-based-fine-grained-inference commit 04b3aaa8a4

Frequently asked questions

npx skillmds@latest add ndpvt-web/profinfer-ebpf-based-fine-grained-inference