Qwen1.5-MoE-A2.7B

View as Markdown

Qwen1.5-MoE is a Mixture-of-Experts variant from Alibaba Cloud that activates only a fraction of parameters per token, enabling efficient training and inference at scale.

Set up NeMo AutoModel with the latest container or follow the installation instructions.

Fine-Tune Qwen1.5-MoE-A2.7B

From the repository root, run:

uv run automodel --nproc-per-node=8 examples/llm_finetune/qwen/qwen1_5_moe_a2_7b_qlora.yaml

Choose a Workflow

GoalStart Here
Quantized low-rank adaptation (QLoRA) - Qwen1.5 MoE A2.7BUse qwen1_5_moe_a2_7b_qlora.yaml.

Model Reference

Model Architecture

PropertyValue
TaskText Generation (MoE)
ArchitectureQwen2MoeForCausalLM
Parameters14.3B total / 2.7B active
Hugging Face OrganizationQwen

Available Models

ModelHF ID
Qwen1.5 MoE A2.7BQwen/Qwen1.5-MoE-A2.7B
Qwen1.5 MoE A2.7B ChatQwen/Qwen1.5-MoE-A2.7B-Chat