Qwen1.5-MoE-A2.7B
Qwen1.5-MoE-A2.7B
Qwen1.5-MoE is a Mixture-of-Experts variant from Alibaba Cloud that activates only a fraction of parameters per token, enabling efficient training and inference at scale.
Set up NeMo AutoModel with the latest container or follow the installation instructions.
Fine-Tune Qwen1.5-MoE-A2.7B
From the repository root, run: