Tao Generate Video Reasoning Annotations

Multi-step video annotation pipeline that turns raw videos into Chain-of-Thought training data — multi-level captions, structured descriptions, and QA pairs (MCQ, binary, open-ended) with reasoning traces, via VLM/LLM distillation.

About this skill

Multi-step video annotation pipeline that turns raw videos into Chain-of-Thought training data — multi-level captions, structured descriptions, and QA pairs (MCQ, binary, open-ended) with reasoning traces, via VLM/LLM distillation.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Purpose
  • 02Pipeline architecture
  • 03Initial consultation
  • 041. Videos
  • 052. Domain — drives prompt selection
  • 063. Anomaly / normal / mixed

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-generate-video-reasoning-annotations

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices