Tensorrt

NVIDIA TensorRT — deep learning inference optimizer. FP16/INT8/INT4 quantization, kernel auto-tuning, layer fusion, and dynamic shapes. Max throughput on NVIDIA GPUs for production inference.

mkurman d9121b1 3.6 KB Updated

File contents

mkurman/zorai/tree/main/skills/scientific-skills/tensorrt commit d9121b17d0

Frequently asked questions

npx skillmds@latest add mkurman/tensorrt