Ling-1T

View as Markdown

Ling-1T has a checked-in NeMo AutoModel recipe for text generation. The Hugging Face configuration declares the BailingMoeV2ForCausalLM architecture.

Fine-Tune Ling-1T

Follow the installation instructions, then run the recipe from the repository root:

automodel examples/llm_finetune/ling/ling_1t_sft.yaml --nproc-per-node 8

Use the Slurm launcher guide for the multi-node run.

Choose a Workflow

GoalStart Here
Run the primary recipe for this modelUse the recipe configuration.
Try another checked-in recipe for this modelUse the alternate recipe.

Configuration

SettingConfiguration
Hardware-
StrategyFSDP2; tp_size=1, pp_size=4, cp_size=1, ep_size=64
Nodes32
FeaturesCheckpointing (enabled=true), Activation checkpointing (activation_checkpointing=true), Transformer Engine, HybridEP (dispatcher=hybridep)
Advancedbfloat16, attention=sdpa, linear=te, experts=torch_mm, dispatcher=hybridep

More Recipes for This Model

Model Reference

Model Architecture

PropertyValue
TaskText generation
ArchitectureBailingMoeV2ForCausalLM

Available Models