Tao Train Action Recognition

Action recognition from video sequences. Supports RGB, optical flow, and joint (multi-stream) input types for classifying temporal actions in video clips.

About this skill

Action recognition from video sequences. Supports RGB, optical flow, and joint (multi-stream) input types for classifying temporal actions in video clips.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Quick Start (docker run)
  • 02Dataclass Schemas
  • 03Train Action Policy
  • 04Training Requirements
  • 05Per-Action Dataset Requirements
  • 06Typical Spec Overrides

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-train-action-recognition

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices