Qwen3-30B-A3B

View as Markdown

Qwen3 MoE is the Mixture-of-Experts variant of the Qwen3 series from Alibaba Cloud, activating a small fraction of parameters per token for efficient large-scale training.

Set up NeMo AutoModel with the latest container or follow the installation instructions.

Fine-Tune Qwen3-30B-A3B

From the repository root, run:

uv run automodel --nproc-per-node=8 examples/llm_finetune/qwen/qwen3_moe_30b_te_deepep.yaml

Choose a Workflow

GoalStart Here
Supervised fine-tuning (SFT) - Qwen3 MoE 30B with TE + DeepEPUse qwen3_moe_30b_te_deepep.yaml.
Low-rank adaptation (LoRA) - Qwen3 MoE 30BUse qwen3_moe_30b_lora.yaml.

Model Reference

Model Architecture

PropertyValue
TaskText Generation (MoE)
ArchitectureQwen3MoeForCausalLM
Parameters30B total / 3B active
Hugging Face OrganizationQwen

Available Models

ModelHF ID
Qwen3 30B A3BQwen/Qwen3-30B-A3B