Nvidia Tensorrt LLM Deployment Review

Use this skill when reviewing TensorRT or TensorRT-LLM deployment artifacts statically — ONNX/PyTorch export pipelines, precision selection (FP16/BF16/INT8/FP8/INT4), calibration cache integrity, dynamic shape profiles, custom plugin loading, engine cache and serialized engine provenance, runtime memory pool sizing. Trigger when the user asks whether a TensorRT build script, calibration pipeline, or trtexec invocation follows NVIDIA's published guidance.

VincentChuWaiChow Updated

File contents

VincentChuWaiChow/vanguard-frontier-agentic/tree/main/skills/nvidia/nvidia-tensorrt-llm-deployment-review commit 01379aa238

Frequently asked questions

npx skillmds@latest add vincentchuwaichow/nvidia-tensorrt-llm-deployment-review