Tao Train Ocrnet

OCRNet for scene text recognition. Recognizes text content from cropped text-region images and supports CTC and attention-based decoders.

About this skill

OCRNet for scene text recognition. Recognizes text content from cropped text-region images and supports CTC and attention-based decoders.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Dataclass Schemas
  • 02Train Action Policy
  • 03Training Requirements
  • 04Per-Action Dataset Requirements
  • 05Checkpoint Selection
  • 06Typical Spec Overrides

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / tao-train-ocrnet

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices