Wan2.1-T2V-1.3B-Diffusers
Wan2.1-T2V-1.3B-Diffusers
Wan 2.1 is a text-to-video diffusion model from Wan AI, trained with flow matching on a large-scale video dataset. It generates high-quality short video clips from text prompts.
Set up NeMo AutoModel with the latest container or follow the installation instructions.
The 1.3B fine-tuning configuration uses 16 data-parallel ranks. Use the launcher guide to run it across the required ranks.
Fine-Tune Wan2.1-T2V-1.3B-Diffusers
From the repository root, run:
Choose a Workflow
Model Reference
Model Architecture
Task
- Text-to-Video (T2V)