Tao Run Inference Service

Start, query, and stop a network-specific TAO inference microservice ({network_arch}-inference-microservice) by delegating container execution to the appropriate platform skill. Handles container image resolution, job-payload JSON construction, and the service registry. Use when the user wants to run inference on a TAO model checkpoint using a microservice container, deploy a TAO inference endpoint, or stop a running inference container.

NVIDIA-TAO Updated

File contents

NVIDIA-TAO/tao-skills-bank/tree/main/skills/applications/tao-run-inference-service commit 63efb4ae50

Frequently asked questions

npx skillmds@latest add nvidia-tao/tao-run-inference-service