tao-run-inference-service

nvidia/tao-run-inference-service · Agent Skill (multi-file)

by NVIDIA · bundle

Published · Last updated


Start, query, and stop a TAO inference microservice for a specific network architecture by delegating container execution to the appropriate platform skill.

SKILL.md

Files

This skill is a package of 11 files. Install with the command above, or download the folder.

Related

  1. tao-launch-workflow · nvidia bundle
    Collects launch inputs and runs preflight checks before executing TAO workflows such as AutoML, training, evaluation, inference, export, TensorRT engine generation, or DEFT jobs on supported platforms.
    2.2k
    repo stars
  2. tao-run-platform · nvidia bundle
    Submit and monitor GPU training jobs on Brev, SLURM, Docker, or Kubernetes using the TAO Execution SDK, with job handles, S3 I/O wrapping, and multi-node distributed training.
    2.2k
    repo stars
  3. gke-app-onboarding · google bundle
    Containerizes applications and deploys them to Google Kubernetes Engine (GKE) with Dockerfiles, manifests, and best practices.
    14.4k
    repo stars
  4. gke-inference · google
    Deploys and optimizes AI/ML inference workloads on GKE, using GPUs, TPUs, and model servers.
    14.4k
    repo stars
  5. deployment-patterns · affaan-m
    Provides deployment strategies, CI/CD pipeline patterns, Docker containerization best practices, health checks, and production readiness guidance for web applications.
    226k
    repo stars
  6. docker · oyi77
    Guides containerizing apps with Docker Compose, optimizing Dockerfiles, and deploying to Kubernetes, with a focus on turning these tasks into a paid service.
    10
    repo stars

Frequently asked questions

How do I install the tao-run-inference-service skill?

Run npx skillmds add nvidia/tao-run-inference-service in your terminal (requires Node.js), paste this page's agent-chat prompt into Claude, Cursor, or any MCP-connected agent, or download the SKILL.md file and copy it into your agent's skills directory.

What does the tao-run-inference-service skill do?

Start, query, and stop a TAO inference microservice for a specific network architecture by delegating container execution to the appropriate platform skill. It is listed under DevOps & Infra, AI & ML, Containers & Kubernetes, Deployment & Release on SkillMD.

Is tao-run-inference-service safe to use?

SkillMD's automated safety review verdict for this skill is CAUTION. Independent scanners report: SkillSpector: PASS, Skill Scanner: PASS. Capability flags: makes network calls, reads secrets. SkillMD never runs a skill's scripts for you; review the SKILL.md before installing.

Which AI agents work with tao-run-inference-service?

This skill is tagged as working with Claude Code, Claude.ai, OpenAI Codex. SKILL.md is an open format, so most agents that read a skills directory can load it too.

Is tao-run-inference-service free to use?

Yes. Installing skills from SkillMD is free. This skill is licensed under Apache-2.

Who published tao-run-inference-service?

NVIDIA (@nvidia) published this skill as a verified publisher. Their other Agent Skills are listed on their SkillMD profile.