gemma-4-31B-it

View as Markdown

gemma-4-31B-it is a multimodal Gemma 4 checkpoint for image-text inputs. NeMo AutoModel provides full-parameter and Low-Rank Adaptation (LoRA) recipes, including tensor-parallel and pipeline-parallel variants.

Set up NeMo AutoModel with the latest container or follow the installation instructions.

Fine-Tune gemma-4-31B-it

From the repository root, run the 8-GPU recipe:

uv run automodel examples/vlm_finetune/gemma4/gemma4_31b.yaml --nproc-per-node 8

Choose a Workflow

GoalStart Here
Fine-tune on MedPix-VQAUse gemma4_31b.yaml.
Fine-tune with LoRAUse gemma4_31b_peft.yaml.
Fine-tune with tensor parallelismUse gemma4_31b_tp4.yaml.
Fine-tune with tensor and pipeline parallelismUse gemma4_31b_tp4_pp2.yaml or the multi-node PP4 recipe.

Model Reference

Model Architecture

PropertyValue
TaskImage-text-to-text
Hugging Face ArchitectureGemma4ForConditionalGeneration
Checkpointgoogle/gemma-4-31B-it