Nvidia Tensorrt LLM Trtllm Serve Config Guide

Generate a source-backed starting `trtllm-serve --config` YAML for basic aggregate single-node PyTorch serving, aligned with checked-in TensorRT-LLM configs and deployment docs. Preserves explicit latency / balanced / throughput objectives. Excludes disaggregated, multi-node, and non-MTP speculative configs.

autohandai Updated

File contents

autohandai/community-skills/tree/main/nvidia-tensorrt-llm-trtllm-serve-config-guide commit 01c875acf2

Frequently asked questions

npx skillmds@latest add autohandai-community-skills/nvidia-tensorrt-llm-trtllm-serve-config-guide