> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.nvidia.com/nemo/automodel/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.nvidia.com/nemo/automodel/_mcp/server.

# NVIDIA-Nemotron-3-Nano-4B-BF16

> Use NVIDIA-Nemotron-3-Nano-4B-BF16 with NeMo AutoModel for language model fine-tuning, with documented checkpoints, runnable recipes, setup, and model reference details.

[NVIDIA Nemotron-H](https://developer.nvidia.com/blog/nemotron-h-reasoning-enabling-throughput-gains-with-no-compromises/) is a hybrid Mamba-2 and transformer architecture that interleaves selective state-space layers with standard attention layers for improved efficiency on long sequences.

Set up NeMo AutoModel with the [latest container](https://catalog.ngc.nvidia.com/orgs/nvidia/containers/nemo-automodel) or follow the [installation instructions](/get-started/installation).

## Fine-Tune NVIDIA-Nemotron-3-Nano-4B-BF16

From the repository root, run:

```bash
uv run automodel examples/llm_finetune/nemotron/nemotron_nano_4b_squad.yaml \
  --nproc-per-node 8
```

## Choose a Workflow

| Goal                                                       | Start Here                                                                                                                                                      |
| ---------------------------------------------------------- | --------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Supervised fine-tuning (SFT) - Nemotron-3-Nano 4B on SQuAD | Use [nemotron\_nano\_4b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_4b_squad.yaml).            |
| Low-rank adaptation (LoRA) - Nemotron-3-Nano 4B on SQuAD   | Use [nemotron\_nano\_4b\_squad\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_4b_squad_peft.yaml). |

## Model Reference

### Model Architecture

| Property                  | Value                                   |
| ------------------------- | --------------------------------------- |
| Task                      | Text Generation                         |
| Architecture              | `NemotronHForCausalLM`                  |
| Parameters                | 4B                                      |
| Hugging Face Organization | [nvidia](https://huggingface.co/nvidia) |

### Available Models

| Model              | HF ID                                                                                                   |
| ------------------ | ------------------------------------------------------------------------------------------------------- |
| Nemotron-3-Nano 4B | [`nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16`](https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16) |

## Related Resources

* [LLM Fine-Tuning Guide](/recipes-e2e-examples/sft-peft)