Tao Train Depth Anything V2

Monocular depth estimation using Metric Depth Anything v2 or Relative Depth Anything architectures. Predicts per-pixel depth from single RGB images.

About this skill

Monocular depth estimation using Metric Depth Anything v2 or Relative Depth Anything architectures. Predicts per-pixel depth from single RGB images.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Train Action Policy
  • 02Workflow
  • 03Prerequisites — data accessibility
  • 04Step 1 — Annotation file
  • 05Step 2 — Pair model_type and dataset_name based on your data
  • 06Step 3 — Write spec yaml from Typical Spec Overrides

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-train-depth-anything-v2

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices