> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.nvidia.com/nemo-platform/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.nvidia.com/nemo-platform/_mcp/server.

# About Fine-Tuning

<a id="ft-about" />

Learn how to fine-tune models by making requests to NVIDIA NeMo Customizer through the API. Fine-tuned models you have created can be deployed using NVIDIA NIMs.

## Fine-Tuning Workflow

<a id="ft-workflow" />

At a high level, the fine-tuning workflow consists of the following steps:

1. [Create a Model Entity](/documentation/customizer-reference/manage-model-entities/overview) pointing to your base model checkpoint (stored as a FileSet).
2. Format a compatible [dataset](/documentation/customizer-reference/tutorials/format-training-dataset).
3. [Create a customization job](/documentation/customizer-reference/manage-customization-jobs) referencing the Model Entity.
4. Monitor the job until it completes.
5. The customization job automatically creates either:

* **LoRA jobs**: An adapter attached to the original Model Entity
* **Full fine-tuning jobs**: A new Model Entity with the customized weights

1. [Deploy the model](/documentation/models-and-inference) using the Deployment Management Service.
2. Move on to [Evaluate the output model](/documentation/evaluate-models).

---

## Model Catalog

Explore the model families and sizes supported by NVIDIA NeMo Customizer.

#### [Llama Models](/documentation/customizer-reference/models/llama)

View the available Llama models in the model catalog.

#### [Llama Nemotron Models](/documentation/customizer-reference/models/llama-nemotron)

View the available Llama Nemotron models from NVIDIA, including Nano and Super variants for efficient and advanced instruction tuning.

#### [Phi Models](/documentation/customizer-reference/models/phi)

View the available Phi models from Microsoft, designed for strong reasoning capabilities with efficient deployment.

#### [GPT-OSS Models](/documentation/customizer-reference/models/gpt-oss)

View the available GPT-OSS models supported for Full SFT customization.

#### [Embedding Models](/documentation/customizer-reference/models/embedding)

View the available embedding models for question-answering and retrieval tasks.

## Task Guides

Perform common fine-tuning tasks.

#### [Manage Customization Jobs](/documentation/customizer-reference/manage-customization-jobs)

Create, list, view, and cancel customization jobs.

#### [Manage Model Entities](/documentation/customizer-reference/manage-model-entities/overview)

Create FileSets and Model Entities to prepare base models for customization.

#### [Manage Datasets](/documentation/get-started/core-concepts/manage-files)

Upload and manage datasets for training.

---

## Tutorials

Follow these tutorials to learn how to accomplish common fine-tuning tasks.

#### [Format Training Datasets](/documentation/customizer-reference/tutorials/format-training-dataset)

Learn how to format datasets for different model types.

<small>
  datasets

   

  chat-models

   

  completion-models
</small>

#### [Start a LoRA Customization Job](tutorials/lora-customization-job.ipynb)

Learn how to start a LoRA customization job using a custom dataset.

<small>
  nemo-customizer
</small>

#### [Start a Full SFT Customization Job](tutorials/sft-customization-job.ipynb)

Learn how to start a SFT customization job using a custom dataset.

<small>
  nemo-customizer
</small>

#### [Distill a Model with Knowledge Distillation](tutorials/distillation-customization-job.ipynb)

Learn how to compress a larger teacher model into a smaller student model.

<small>
  nemo-customizer

   

  knowledge-distillation
</small>

#### [Check Customization Job Metrics](/documentation/customizer-reference/tutorials/metrics)

Learn how to check job metrics using MLFlow or Weights & Biases.

<small>
  nemo-customizer

   

  mlflow

   

  wandb
</small>

#### [Optimize Tokens per GPU](tutorials/optimize-throughput.ipynb)

Learn how to optimize the token-per-GPU throughput for a LoRA optimization job.

<small>
  nemo-customizer

   

  wandb

   

  sequence-packing
</small>

---

## References

#### [Hyperparameters](/documentation/customizer-reference/manage-customization-jobs/training-configuration)

View the available hyperparameters and their valid options that you can set when creating a customization job.

#### [Customizer API](/documentation/reference/api-reference)

View the OpenAPI specification for Customizer.

#### [Troubleshoot Failed Jobs](/documentation/reference/troubleshooting/customizer)

View troubleshooting tips for failed jobs.