Nemo Automodel Distributed Training

Guide for selecting and configuring distributed training strategies in NeMo AutoModel, including FSDP2, Megatron FSDP, DDP, and parallelism settings.

About this skill

Guide for selecting and configuring distributed training strategies in NeMo AutoModel, including FSDP2, Megatron FSDP, DDP, and parallelism settings.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Purpose
  • 02Instructions
  • 03Examples
  • 04TP plus PP for a large multi-node model
  • 05MoE expert parallelism
  • 06MegatronFSDP limitations

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / nemo-automodel-distributed-training

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices