Qwen3-Next-80B-A3B-Instruct
Qwen3-Next-80B-A3B-Instruct
Qwen3-Next is an advanced MoE language model from Alibaba Cloud’s Qwen team designed for high-throughput inference with large total parameter counts and efficient per-token activation.
Set up NeMo AutoModel with the latest container or follow the installation instructions.
Fine-Tune Qwen3-Next-80B-A3B-Instruct
This recipe was validated on 4 nodes x 8 GPUs (32 H100s). See the Launcher Guide for multi-node setup.
From the repository root, run: