Physicsnemo Shard Tensor

Official NVIDIA-authored guidance for PhysicsNeMo ShardTensor domain parallelism — integrate domain parallelism into training/inference scripts (new or existing) with DDP or FSDP2, write and register shard patches to enable new…

About this skill

Official NVIDIA-authored guidance for PhysicsNeMo ShardTensor domain parallelism — integrate domain parallelism into training/inference scripts (new or existing) with DDP or FSDP2, write and register shard patches to enable new layers/ops, and bootstrap multi-GPU correctness tests.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01When NOT to use
  • 02The core promise: the model does not change
  • 03Mesh and data setup (every script)
  • 04Choosing the data-parallel wrapper
  • 05torch.compile with ShardTensor
  • 06Debugging pitfalls (each of these cost real time — check them first)

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / physicsnemo-shard-tensor

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices