Tensorrt Optimization

Compile a trained model into a fast, GPU-specific TensorRT engine by controlling precision, defining dynamic shape profiles, and proving the fused engine kept its accuracy. Use when a PyTorch or ONNX model must reach a hardware latency floor that eager execution cannot.

Amey-Thakur 80d9fcc 3.5 KB Updated

File contents

Amey-Thakur/AI-SKILLS/tree/main/skills/gpu-ai-infrastructure/tensorrt-optimization commit 80d9fcc878

Frequently asked questions

npx skillmds@latest add amey-thakur/tensorrt-optimization