Tao Train Bevfusion

BEVFusion for multi-sensor 3D object detection. Fuses LiDAR point clouds and camera images in bird's-eye-view (BEV) space, used in autonomous driving for robust 3D perception.

About this skill

BEVFusion for multi-sensor 3D object detection. Fuses LiDAR point clouds and camera images in bird's-eye-view (BEV) space, used in autonomous driving for robust 3D perception.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Dataclass Schemas
  • 02Train Action Policy
  • 03Training Requirements
  • 04Per-Action Dataset Requirements
  • 05Typical Spec Overrides
  • 06Eval Dataset

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-train-bevfusion

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices