Techtide Nvidia Tensorrt LLM Deployment Review

Use this skill when reviewing TensorRT or TensorRT-LLM deployment artifacts statically - ONNX/PyTorch export pipelines, precision selection (FP16/BF16/INT8/FP8/INT4), calibration cache integrity, dynamic shape profiles, custom plugin loading, engine cache and serialized engine provenance, runtime memory pool sizing. Trigger when the user asks whether a TensorRT build script, calibration pipeline, or trtexec invocation follows NVIDIA's published guidance.

TechTideOhio 5cd01f0 2 files · 4.5 KB Updated

File contents

TechTideOhio/techtide-harness-kit/tree/main/skills/nvidia/techtide-nvidia-tensorrt-llm-deployment-review commit 5cd01f0487

Frequently asked questions

npx skillmds@latest add techtideohio/techtide-nvidia-tensorrt-llm-deployment-review