Vllm Installer

This skill should be used when users need to install, configure, debug, or run vLLM inference server on NVIDIA GPUs (especially B200/H100/A100). It covers installation from PyPI or source, dependency management, environment setup, common error diagnosis and fixes, tensor parallelism configuration, and server startup/testing. The skill automatically checks for LSSD mount status and DeepEP installation for MoE models.

yangwhale a04551e 5 files · 38.3 KB Updated

File contents

yangwhale/gpu-tpu-pedia/tree/main/VibeCoding/claude-code/skills/vllm-installer commit a04551eff9

Frequently asked questions

npx skillmds@latest add yangwhale/vllm-installer