Tao Run Inference Service

Start, query, and stop a network-specific TAO inference microservice ({networkarch}-inference-microservice) by delegating container execution to the appropriate platform skill. Handles container image resolution, job-payload JSON…

About this skill

Start, query, and stop a network-specific TAO inference microservice ({networkarch}-inference-microservice) by delegating container execution to the appropriate platform skill. Handles container image resolution, job-payload JSON construction, and the service registry.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Instructions
  • 02Secrets rule (applies to every generated code block in this skill)
  • 031. What to collect from the user
  • 042. Image resolution
  • 053. Environment variables (no callbacks)
  • 064. Executing across platforms

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-run-inference-service

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices