Model CoverageLarge Language ModelsNVIDIANVIDIA-Nemotron-3-Super-120B-A12B-BF16

NVIDIA-Nemotron-3-Super-120B-A12B-BF16

View as Markdown

NVIDIA-Nemotron-3-Super-120B-A12B-BF16 has a checked-in NeMo AutoModel recipe for text generation. The Hugging Face configuration declares the NemotronHForCausalLM architecture.

Fine-Tune NVIDIA-Nemotron-3-Super-120B-A12B-BF16

Follow the installation instructions, then review the checked-in recipe and its declared topology below before choosing a local or cluster launcher.

Use the Slurm launcher guide for the multi-node run.

Choose a Workflow

GoalStart Here
Run the primary recipe for this modelUse the recipe configuration.
Try another checked-in recipe for this modelUse the alternate recipe.

Configuration

SettingConfiguration
Hardware-
StrategyFSDP2; tp_size=1, cp_size=1, ep_size=32
Nodes4
FeaturesCheckpointing (enabled=true)
Advanced-

More Recipes for This Model

Model Reference

Model Architecture

PropertyValue
TaskText generation
ArchitectureNemotronHForCausalLM

Available Models