Qwen3.5-35B-A3B

View as Markdown

Qwen3.5-35B-A3B has a checked-in NeMo AutoModel recipe for text generation. The Hugging Face configuration declares the Qwen3_5MoeForConditionalGeneration architecture.

Fine-Tune Qwen3.5-35B-A3B

Follow the installation instructions, then run the recipe from the repository root:

automodel examples/llm_benchmark/qwen/qwen3.5_moe_lora.yaml --nproc-per-node 8

Choose a Workflow

GoalStart Here
Run the primary recipe for this modelUse the recipe configuration.
Try another checked-in recipe for this modelUse the alternate recipe.

Configuration

SettingConfiguration
Hardware-
StrategyFSDP2; tp_size=1, pp_size=1, cp_size=1, ep_size=8
Nodes1
FeaturesActivation checkpointing (activation_checkpointing=true), LoRA, Transformer Engine, DeepEP (dispatcher=deepep)
Advancedattention=te, linear=te, experts=gmm, dispatcher=deepep

More Recipes for This Model

Model Reference

Model Architecture

PropertyValue
TaskText generation
ArchitectureQwen3_5MoeForConditionalGeneration

Available Models