Qwen2.5-0.5B

View as Markdown

Qwen2.5-0.5B has a checked-in NeMo AutoModel recipe for text generation. The Hugging Face configuration declares the Qwen2ForCausalLM architecture.

Fine-Tune Qwen2.5-0.5B

Follow the installation instructions, then run the recipe from the repository root:

automodel examples/llm_finetune/qwen/qwen2_5_0p5b_packed_cp2.yaml --nproc-per-node 2

Choose a Workflow

GoalStart Here
Run the primary recipe for this modelUse the recipe configuration.
Try another checked-in recipe for this modelUse the alternate recipe.

Configuration

SettingConfiguration
Hardware-
StrategyFSDP2; tp_size=1, pp_size=1, cp_size=2, ep_size=1
Nodes1
FeaturesPacked sequences (packed_sequence_size=1024), Transformer Engine
Advancedbf16, attention=te, linear=torch

More Recipes for This Model

Model Reference

Model Architecture

PropertyValue
TaskText generation
ArchitectureQwen2ForCausalLM

Available Models