> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.nvidia.com/nemo/automodel/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.nvidia.com/nemo/automodel/_mcp/server.

# Model Support Log

> A reverse-chronological log of model support added to NeMo AutoModel.

A reverse-chronological log of model support added to NeMo AutoModel.

See the [LLM](/model-coverage/large-language-models/overview) / [VLM](/model-coverage/vision-language-models/overview) / [Multimodal](/model-coverage/multimodal/overview) / [Omni](/model-coverage/omni/overview) / [dLLM](/model-coverage/dllm/overview) / [Diffusion](/model-coverage/diffusion/overview) / [Embedding](/model-coverage/embedding-models/overview) / [Reranker](/model-coverage/reranking-models/overview) pages for the full architecture listings.

Browse release tables on the [LLM](/model-coverage/large-language-models/overview#model-support-log), [VLM](/model-coverage/vision-language-models/overview#model-support-log), [Multimodal](/model-coverage/multimodal/overview#model-support-log), [Omni](/model-coverage/omni/overview#model-support-log), [dLLM](/model-coverage/dllm/overview#model-support-log), [Diffusion](/model-coverage/diffusion/overview#model-support-log), [Embedding](/model-coverage/embedding-models/overview#model-support-log), and [Reranking](/model-coverage/reranking-models/overview#model-support-log) overview pages.

| Date       | Type            | Model                                                                                                                                                                                                                                                                                                                 |
| :--------- | :-------------- | :-------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| 2026-08-08 | LLM             | [NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16](https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16) ([nemotron\_nano\_v3\_5\_lightning\_hellaswag\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_v3_5_lightning_hellaswag_peft.yaml)) |
| 2026-08-07 | LLM             | [Kimi-Linear-48B-A3B-Instruct](/model-coverage/large-language-models/moonshotai/kimi-linear) ([kimi\_linear\_48b\_a3b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/kimi/kimi_linear_48b_a3b_hellaswag.yaml))                                                             |
| 2026-07-30 | dLLM            | [Qwen3-8B](/model-coverage/dllm/qwen/qwen3-idlm) ([qwen3\_8b\_idlm.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/qwen3_8b_idlm.yaml))                                                                                                                                                    |
| 2026-07-30 | Encoder-Decoder | [T5-small](/model-coverage/large-language-models/google-t5/t5) ([t5\_small\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/t5/t5_small_squad.yaml))                                                                                                                             |
| 2026-07-30 | VLM             | [Inkling-Small](/model-coverage/vision-language-models/thinkingmachines/inkling) ([Inkling\_small\_medpix\_ep64.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/inkling/Inkling_small_medpix_ep64.yaml))                                                                               |
| 2026-07-29 | LLM             | [Kimi-K3](/model-coverage/large-language-models/moonshotai/kimi-k3) ([k3\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/kimi/k3_hellaswag.yaml))                                                                                                                           |
| 2026-07-29 | VLM             | [Qwen3.5-122B-A10B](https://huggingface.co/Qwen/Qwen3.5-122B-A10B) ([qwen3\_5\_122b\_128k\_ep8cp32.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5_moe/qwen3_5_122b_128k_ep8cp32.yaml))                                                                                        |
| 2026-07-28 | Diffusion       | [LTX-2.3-Diffusers](/model-coverage/diffusion/diffusers/ltx-2-3) ([ltx2\_3\_t2v\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/ltx2_3_t2v_flow.yaml))                                                                                                                     |
| 2026-07-27 | Diffusion       | [Qwen-Image-Edit-2511](/model-coverage/diffusion/qwen/qwen-image) ([qwen\_image\_edit\_2511\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/qwen_image_edit_2511_flow.yaml))                                                                                               |
| 2026-07-22 | LLM             | [Gemma-4-31B](https://huggingface.co/google/gemma-4-31B) ([gemma4\_31b\_base\_coderforge\_cp8\_64k\_1e5\_800steps.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/long_context_validation/gemma4_31B/gemma4_31b_base_coderforge_cp8_64k_1e5_800steps.yaml))                                         |
| 2026-07-22 | LLM             | [Laguna-S-2.1](/model-coverage/large-language-models/poolside/laguna) ([laguna\_s\_2p1\_hellaswag\_ep16.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/laguna/laguna_s_2p1_hellaswag_ep16.yaml))                                                                                      |
| 2026-07-17 | VLM             | [Inkling](/model-coverage/vision-language-models/thinkingmachines/inkling) ([inkling\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/inkling/inkling_medpix.yaml))                                                                                                             |
| 2026-06-27 | Embedding       | [Llama-nemotron-embed-vl-1b-v2](https://huggingface.co/nvidia/llama-nemotron-embed-vl-1b-v2) ([nemotron\_vl\_1b\_example.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/retrieval/bi_encoder/nemotron_vl_1b/nemotron_vl_1b_example.yaml))                                                          |
| 2026-06-21 | LLM             | [GLM-5.2](/model-coverage/large-language-models/thudm/glm-5-moe-dsa) ([glm\_5.2\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_5.2_hellaswag_pp.yaml))                                                                                                         |
| 2026-06-18 | LLM             | [Qwen2.5-0.5B](https://huggingface.co/Qwen/Qwen2.5-0.5B) ([qwen25\_magi\_prefix\_tree\_rollouts.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen25_magi_prefix_tree_rollouts.yaml))                                                                                           |
| 2026-06-12 | VLM             | [MiniMax-M3](/model-coverage/vision-language-models/minimax/minimax-m3) ([minimax\_m3\_vl\_lora\_pp4ep8\_8node.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/minimax_m3/minimax_m3_vl_lora_pp4ep8_8node.yaml))                                                                       |
| 2026-06-10 | dLLM            | [Diffusiongemma-26B-A4B-it](/model-coverage/dllm/google/diffusiongemma) ([diffusion\_gemma\_lora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/diffusion_gemma_lora.yaml))                                                                                                               |
| 2026-06-09 | Diffusion       | [FLUX.2-dev](/model-coverage/diffusion/black-forest-labs/flux-2-dev) ([flux2\_t2i\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/flux2_t2i_flow.yaml))                                                                                                                    |
| 2026-06-08 | Diffusion       | [Wan2.2-T2V-A14B-Diffusers](/model-coverage/diffusion/wan-ai/wan-2-2-t2v-a14b) ([wan2\_2\_t2v\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/wan2_2_t2v_flow.yaml))                                                                                                       |
| 2026-06-08 | LLM             | [Falcon-H1-0.5B-Instruct](https://huggingface.co/tiiuae/Falcon-H1-0.5B-Instruct) ([falcon\_h1\_0p5b\_instruct\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/falcon/falcon_h1_0p5b_instruct_squad.yaml))                                                                       |
| 2026-06-08 | LLM             | [Falcon-H1-1.5B-Deep-Instruct](https://huggingface.co/tiiuae/Falcon-H1-1.5B-Deep-Instruct) ([falcon\_h1\_1p5b\_instruct\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/falcon/falcon_h1_1p5b_instruct_squad.yaml))                                                             |
| 2026-06-08 | LLM             | [Falcon-H1-34B-Instruct](https://huggingface.co/tiiuae/Falcon-H1-34B-Instruct) ([falcon\_h1\_34b\_instruct\_squad\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/falcon/falcon_h1_34b_instruct_squad_peft.yaml))                                                                |
| 2026-06-08 | LLM             | [Falcon-H1-7B-Instruct](https://huggingface.co/tiiuae/Falcon-H1-7B-Instruct) ([falcon\_h1\_7b\_instruct\_hellaswag\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/falcon/falcon_h1_7b_instruct_hellaswag_peft.yaml))                                                            |
| 2026-06-08 | Multimodal      | [BAGEL-7B-MoT](/model-coverage/multimodal/bytedance-seed/bagel) ([bagel\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/multimodal_finetune/bagel/bagel_sft.yaml))                                                                                                                             |
| 2026-06-04 | LLM             | [Gemma-4-12B](https://huggingface.co/google/gemma-4-12B) ([gemma\_4\_12b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gemma/gemma_4_12b_hellaswag.yaml))                                                                                                                 |
| 2026-06-04 | LLM             | [NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16](https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16) ([nemotron\_ultra\_v3\_hellaswag\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_ultra_v3_hellaswag_peft.yaml))                                 |
| 2026-06-03 | dLLM            | [LLaDA2.1-mini](/model-coverage/dllm/inclusionai/llada2) ([llada2\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/llada2_sft.yaml))                                                                                                                                                   |
| 2026-06-03 | dLLM            | [Qwen3-4B-DFlash-b16](/model-coverage/dllm/z-lab/dflash) ([dflash\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/dflash_sft.yaml))                                                                                                                                                   |
| 2026-06-02 | LLM             | [MiniCPM5-1B](/model-coverage/large-language-models/openbmb/minicpm) ([minicpm5\_1b\_hellaswag\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/minicpm/minicpm5_1b_hellaswag_peft.yaml))                                                                                         |
| 2026-05-29 | Omni            | [Qwen2.5-Omni-3B](/model-coverage/omni/qwen/qwen2-5-omni) ([ami\_sft\_3b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/audio_finetune/qwen2_5_omni_asr/ami_sft_3b.yaml))                                                                                                                          |
| 2026-05-29 | Omni            | [Qwen2.5-Omni-7B](/model-coverage/omni/qwen/qwen2-5-omni) ([ami\_sft\_7b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/audio_finetune/qwen2_5_omni_asr/ami_sft_7b.yaml))                                                                                                                          |
| 2026-05-29 | VLM             | [Step-3.7-Flash](/model-coverage/vision-language-models/stepfun-ai/step-3-7-flash) ([step3p7\_medpix\_200b\_ep32pp4.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/stepfun/step3p7_medpix_200b_ep32pp4.yaml))                                                                         |
| 2026-05-28 | LLM             | [Qwen2.5-3B](https://huggingface.co/Qwen/Qwen2.5-3B) ([qwen2\_5\_3b\_function\_calling.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/agent/qwen2_5_3b_function_calling.yaml))                                                                                                        |
| 2026-05-28 | LLM             | [Hy-MT2-30B-A3B](/model-coverage/large-language-models/tencent/hy-mt2) ([hy\_mt2\_30b\_a3b\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/hy_mt2/hy_mt2_30b_a3b_sft.yaml))                                                                                                       |
| 2026-05-23 | dLLM            | [Nemotron-Labs-Diffusion-8B-Base](/model-coverage/dllm/nvidia/nemotron-labs-diffusion) ([nemotron\_labs\_diffusion\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/nemotron_labs_diffusion_sft.yaml))                                                                                 |
| 2026-05-22 | LLM             | [DeepSeek-V4-Pro](/model-coverage/large-language-models/deepseek-ai/deepseek-v4-pro) ([deepseek\_v4\_pro\_hellaswag\_all\_tilelang\_pp8\_ep64\_20steps.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/deepseek_v4/deepseek_v4_pro_hellaswag_all_tilelang_pp8_ep64_20steps.yaml))      |
| 2026-05-21 | Embedding       | [Ministral-3-3B-Instruct-2512-BF16](https://huggingface.co/mistralai/Ministral-3-3B-Instruct-2512-BF16) ([ministral3\_3b\_instruct.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/retrieval/bi_encoder/ministral3_3b_instruct.yaml))                                                               |
| 2026-05-20 | LLM             | [Ling-1T](/model-coverage/large-language-models/inclusionai/ling-2-0) ([ling\_1t\_lora\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/ling/ling_1t_lora_pp.yaml))                                                                                                                 |
| 2026-05-20 | LLM             | [Ling-flash-2.0](/model-coverage/large-language-models/inclusionai/ling-2-0) ([ling\_flash\_2\_0\_lora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/ling/ling_flash_2_0_lora.yaml))                                                                                                 |
| 2026-05-20 | LLM             | [Ling-mini-2.0](/model-coverage/large-language-models/inclusionai/ling-2-0) ([ling\_mini\_2\_0\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/ling/ling_mini_2_0_hellaswag.yaml))                                                                                          |
| 2026-05-17 | LLM             | [ERNIE-4.5-0.3B-PT](/model-coverage/large-language-models/baidu/ernie-4-5) ([ernie4\_5\_0p3b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/ernie4_5/ernie4_5_0p3b_hellaswag.yaml))                                                                                        |
| 2026-05-17 | LLM             | [ERNIE-4.5-21B-A3B-PT](/model-coverage/large-language-models/baidu/ernie-4-5) ([ernie4\_5\_21b\_a3b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/ernie4_5/ernie4_5_21b_a3b_hellaswag.yaml))                                                                              |
| 2026-05-17 | LLM             | [MiMo-V2-Flash](/model-coverage/large-language-models/xiaomimimo/mimo-v2-flash) ([mimo\_v2\_flash\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mimo_v2_flash/mimo_v2_flash_hellaswag.yaml))                                                                              |
| 2026-04-29 | LLM             | [Hy3-preview](/model-coverage/large-language-models/tencent/hy3-preview) ([hy3\_preview\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/hy_v3/hy3_preview_deepep.yaml))                                                                                                        |
| 2026-04-29 | VLM             | [Mistral-Medium-3.5-128B](https://huggingface.co/mistralai/Mistral-Medium-3.5-128B) ([mistral3p5\_128b\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/mistral3p5/mistral3p5_128b_medpix.yaml))                                                                                |
| 2026-04-27 | LLM             | [DeepSeek-V4-Flash](/model-coverage/large-language-models/deepseek-ai/deepseek-v4-flash) ([deepseek\_v4\_flash\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/deepseek_v4/deepseek_v4_flash_hellaswag.yaml))                                                               |
| 2026-04-27 | VLM             | [Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16](https://huggingface.co/nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16) ([nemotron\_omni\_cord\_v2.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/nemotron_omni/nemotron_omni_cord_v2.yaml))                                         |
| 2026-04-22 | VLM             | [Qwen3.6-27B](/model-coverage/vision-language-models/qwen/qwen3-6-vl) ([qwen3\_6\_27b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5/qwen3_6_27b.yaml))                                                                                                                       |
| 2026-04-21 | VLM             | [Qwen3.5-27B](https://huggingface.co/Qwen/Qwen3.5-27B) ([qwen3\_5\_27b\_tp4pp4.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5/qwen3_5_27b_tp4pp4.yaml))                                                                                                                       |
| 2026-04-20 | Diffusion       | [Qwen-Image](/model-coverage/diffusion/qwen/qwen-image) ([qwen\_image\_t2i\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/qwen_image_t2i_flow.yaml))                                                                                                                      |
| 2026-04-17 | VLM             | [LLaVA-OneVision-1.5-4B-Instruct](/model-coverage/vision-language-models/lmms-lab/llava-onevision) ([llava\_ov\_1\_5\_4b\_finetune.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/llava_onevision/llava_ov_1_5_4b_finetune.yaml))                                                     |
| 2026-04-17 | VLM             | [LLaVA-OneVision-1.5-8B-Instruct](/model-coverage/vision-language-models/lmms-lab/llava-onevision) ([llava\_ov\_1\_5\_8b\_lora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/llava_onevision/llava_ov_1_5_8b_lora.yaml))                                                             |
| 2026-04-16 | VLM             | [Qwen3.6-35B-A3B](/model-coverage/vision-language-models/qwen/qwen3-6-vl) ([qwen3\_6\_35b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5_moe/qwen3_6_35b.yaml))                                                                                                               |
| 2026-04-15 | LLM             | [Llama-3.1-8B-Instruct](https://huggingface.co/meta-llama/Llama-3.1-8B-Instruct) ([customizer\_llama\_3\_1\_8b\_full\_sft\_tp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_1/customizer_llama_3_1_8b_full_sft_tp.yaml))                                                      |
| 2026-04-15 | LLM             | [Llama-3.2-1B-Instruct](https://huggingface.co/meta-llama/Llama-3.2-1B-Instruct) ([customizer\_llama\_3\_2\_1b\_full\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_2/customizer_llama_3_2_1b_full_sft.yaml))                                                             |
| 2026-04-12 | LLM             | [DeepSeek-V3.2](/model-coverage/large-language-models/deepseek-ai/deepseek-v3) ([dsv32\_lora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_benchmark/deepseek/dsv32_lora.yaml))                                                                                                               |
| 2026-04-12 | LLM             | [Mistral-Small-4-119B-2603](https://huggingface.co/mistralai/Mistral-Small-4-119B-2603) ([mistral\_small\_4\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_benchmark/mistral/mistral_small_4_te_deepep.yaml))                                                                      |
| 2026-04-12 | LLM             | [Kimi-K2-Base](https://huggingface.co/moonshotai/Kimi-K2-Base) ([kimi\_k2\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_benchmark/kimi/kimi_k2_te_deepep.yaml))                                                                                                                   |
| 2026-04-12 | LLM             | [Qwen3-235B-A22B](/model-coverage/large-language-models/qwen/qwen3-moe) ([qwen3\_moe\_235b\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_benchmark/qwen/qwen3_moe_235b_te_deepep.yaml))                                                                                           |
| 2026-04-12 | LLM             | [Qwen3.5-35B-A3B](https://huggingface.co/Qwen/Qwen3.5-35B-A3B) ([qwen3.5\_moe\_lora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_benchmark/qwen/qwen3.5_moe_lora.yaml))                                                                                                                      |
| 2026-04-11 | LLM             | [MiniMax-M2.7](/model-coverage/large-language-models/minimax/minimax-m2) ([minimax\_m2.7\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/minimax_m2/minimax_m2.7_hellaswag_pp.yaml))                                                                                    |
| 2026-04-07 | LLM             | [GLM-5.1](/model-coverage/large-language-models/thudm/glm-5-moe-dsa) ([glm\_5.1\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_5.1_hellaswag_pp.yaml))                                                                                                         |
| 2026-04-06 | LLM             | [Ministral-3-3B-Instruct-2512](/model-coverage/large-language-models/mistralai/ministral3-devstral) ([ministral3\_3b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mistral/ministral3_3b_squad.yaml))                                                                         |
| 2026-04-06 | LLM             | [Llama-3.1-Nemotron-Nano-8B-v1](https://huggingface.co/nvidia/Llama-3.1-Nemotron-Nano-8B-v1) ([nemotron\_nano\_8b\_v1\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_8b_v1_squad.yaml))                                                                 |
| 2026-04-04 | dLLM            | [LLaDA-8B-Base](/model-coverage/dllm/gsai-ml/llada) ([llada\_sft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/dllm_sft/llada_sft.yaml))                                                                                                                                                          |
| 2026-04-03 | LLM             | [Qwen2.5-0.5B-Instruct](https://huggingface.co/Qwen/Qwen2.5-0.5B-Instruct) ([qwen2\_5\_0p5b\_instruct\_fineproofs\_chat.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen2_5_0p5b_instruct_fineproofs_chat.yaml))                                                              |
| 2026-04-02 | VLM             | [Gemma-4-26B-A4B-it](/model-coverage/vision-language-models/google/gemma-4) ([gemma4\_26b\_a4b\_moe.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma4/gemma4_26b_a4b_moe.yaml))                                                                                                   |
| 2026-04-02 | VLM             | [Gemma-4-31B-it](/model-coverage/vision-language-models/google/gemma-4) ([gemma4\_31b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma4/gemma4_31b.yaml))                                                                                                                         |
| 2026-04-02 | VLM             | [Gemma-4-E2B-it](/model-coverage/vision-language-models/google/gemma-4) ([gemma4\_2b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma4/gemma4_2b.yaml))                                                                                                                           |
| 2026-04-02 | VLM             | [Gemma-4-E4B-it](/model-coverage/vision-language-models/google/gemma-4) ([gemma4\_4b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma4/gemma4_4b.yaml))                                                                                                                           |
| 2026-03-30 | Embedding       | [Llama-3.1-8B](/model-coverage/embedding-models/nvidia/llama-bidirectional) ([llama\_embed\_nemotron\_8b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/retrieval/bi_encoder/llama_embed_nemotron_8b/llama_embed_nemotron_8b.yaml))                                                                |
| 2026-03-30 | Embedding       | [Llama-3.2-1B](/model-coverage/embedding-models/nvidia/llama-bidirectional) ([llama3\_2\_1b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/retrieval/bi_encoder/llama3_2_1b.yaml))                                                                                                                 |
| 2026-03-30 | LLM             | [NVIDIA-Nemotron-3-Nano-4B-BF16](/model-coverage/large-language-models/nvidia/nemotron-h) ([nemotron\_nano\_4b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_4b_squad.yaml))                                                                           |
| 2026-03-30 | Reranking       | [Llama-3.2-1B](/model-coverage/reranking-models/nvidia/llama-bidirectional) ([llama3\_2\_1b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/retrieval/cross_encoder/llama3_2_1b.yaml))                                                                                                              |
| 2026-03-26 | LLM             | [Qwen3-30B-A3B-Base](https://huggingface.co/Qwen/Qwen3-30B-A3B-Base) ([qwen3\_moe\_30b\_ep8\_flashoptim.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/convergence/tulu3/models/qwen3-moe-30b/qwen3_moe_30b_ep8_flashoptim.yaml))                                                                  |
| 2026-03-26 | LLM             | [Qwen3-4B-Base](https://huggingface.co/Qwen/Qwen3-4B-Base) ([qwen3\_4b\_cp1\_flashoptim.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/convergence/tulu3/models/qwen3-4b/qwen3_4b_cp1_flashoptim.yaml))                                                                                            |
| 2026-03-16 | VLM             | [Mistral-Small-4-119B-2603](/model-coverage/vision-language-models/mistralai/mistral-small-4) ([mistral4\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/mistral4/mistral4_medpix.yaml))                                                                                       |
| 2026-03-11 | LLM             | [NVIDIA-Nemotron-3-Super-120B-A12B-BF16](https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16) ([nemotron\_super\_v3\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_super_v3_hellaswag.yaml))                                            |
| 2026-03-11 | LLM             | [GLM-5](/model-coverage/large-language-models/thudm/glm-5-moe-dsa) ([glm\_5\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_5_hellaswag_pp.yaml))                                                                                                               |
| 2026-03-09 | LLM             | [Qwen3-30B-A3B-Thinking-2507](https://huggingface.co/Qwen/Qwen3-30B-A3B-Thinking-2507) ([qwen3\_moe\_30b\_te\_chat\_thd.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen3_moe_30b_te_chat_thd.yaml))                                                                          |
| 2026-03-03 | Diffusion       | [FLUX.1-dev](/model-coverage/diffusion/black-forest-labs/flux-1-dev) ([flux\_t2i\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/flux_t2i_flow.yaml))                                                                                                                      |
| 2026-03-03 | Diffusion       | [HunyuanVideo-1.5-Diffusers-720p\_t2v](/model-coverage/diffusion/hunyuanvideo-community/hunyuanvideo-1-5) ([hunyuan\_t2v\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/hunyuan_t2v_flow.yaml))                                                                           |
| 2026-03-03 | Diffusion       | [Wan2.1-T2V-1.3B-Diffusers](/model-coverage/diffusion/wan-ai/wan-2-1-t2v) ([wan2\_1\_t2v\_flow\_multinode.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/wan2_1_t2v_flow_multinode.yaml))                                                                                       |
| 2026-03-03 | Diffusion       | [Wan2.1-T2V-14B-Diffusers](https://huggingface.co/Wan-AI/Wan2.1-T2V-14B-Diffusers) ([wan2\_1\_t2v\_flow.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/diffusion/finetune/wan2_1_t2v_flow.yaml))                                                                                                   |
| 2026-03-02 | VLM             | [Qwen3.5-4B](https://huggingface.co/Qwen/Qwen3.5-4B) ([qwen3\_5\_4b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5/qwen3_5_4b.yaml))                                                                                                                                          |
| 2026-03-02 | VLM             | [Qwen3.5-9B](https://huggingface.co/Qwen/Qwen3.5-9B) ([qwen3\_5\_9b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5/qwen3_5_9b.yaml))                                                                                                                                          |
| 2026-02-24 | VLM             | [Qwen3.5-35B-A3B](https://huggingface.co/Qwen/Qwen3.5-35B-A3B) ([qwen3\_5\_35b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5_moe/qwen3_5_35b.yaml))                                                                                                                          |
| 2026-02-15 | VLM             | [Qwen3.5-397B-A17B](https://huggingface.co/Qwen/Qwen3.5-397B-A17B) ([qwen3\_5\_moe\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3_5_moe/qwen3_5_moe_medpix.yaml))                                                                                                       |
| 2026-02-13 | LLM             | [MiniMax-M2.5](/model-coverage/large-language-models/minimax/minimax-m2) ([minimax\_m2.5\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/minimax_m2/minimax_m2.5_hellaswag_pp.yaml))                                                                                    |
| 2026-02-11 | LLM             | [GLM-4.7-Flash](/model-coverage/large-language-models/thudm/glm-4-moe-glm-4-5-glm-4-7) ([glm\_4.7\_flash\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_4.7_flash_te_deepep.yaml))                                                                                |
| 2026-02-08 | LLM             | [MiniMax-M2.1](/model-coverage/large-language-models/minimax/minimax-m2) ([minimax\_m2.1\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/minimax_m2/minimax_m2.1_hellaswag_pp.yaml))                                                                                    |
| 2026-02-06 | VLM             | [Qwen3-VL-235B-A22B-Instruct](https://huggingface.co/Qwen/Qwen3-VL-235B-A22B-Instruct) ([qwen3\_vl\_moe\_235b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3/qwen3_vl_moe_235b.yaml))                                                                                           |
| 2026-02-04 | LLM             | [Step-3.5-Flash](/model-coverage/large-language-models/stepfun-ai/step-3-5) ([step\_3.5\_flash\_hellaswag\_pp.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/stepfun/step_3.5_flash_hellaswag_pp.yaml))                                                                               |
| 2026-01-31 | VLM             | [Kimi-K2.5](https://huggingface.co/moonshotai/Kimi-K2.5) ([kimi25vl\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/kimi/kimi25vl_medpix.yaml))                                                                                                                                |
| 2026-01-30 | VLM             | [Kimi-VL-A3B-Instruct](/model-coverage/vision-language-models/moonshotai/kimi-vl) ([kimi2vl\_cordv2.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/kimi/kimi2vl_cordv2.yaml))                                                                                                         |
| 2026-01-29 | LLM             | [NVIDIA-Nemotron-3-Nano-30B-A3B-BF16](/model-coverage/large-language-models/nvidia/nemotron-h) ([nemotron\_nano\_v3\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_v3_hellaswag.yaml))                                                              |
| 2026-01-27 | LLM             | [GLM-4.7](/model-coverage/large-language-models/thudm/glm-4-moe-glm-4-5-glm-4-7) ([glm\_4.7\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_4.7_te_deepep.yaml))                                                                                                   |
| 2026-01-12 | LLM             | [Nemotron-Flash-1B](/model-coverage/large-language-models/nvidia/nemotron-flash) ([nemotron\_flash\_1b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron_flash/nemotron_flash_1b_squad.yaml))                                                                            |
| 2026-01-12 | VLM             | [NVIDIA-Nemotron-Parse-v1.1](/model-coverage/vision-language-models/nvidia/nemotron-parse) ([nemotron\_parse\_v1\_1.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/nemotron/nemotron_parse_v1_1.yaml))                                                                                |
| 2026-01-08 | LLM             | [Qwen1.5-MoE-A2.7B](/model-coverage/large-language-models/qwen/qwen2-moe) ([qwen1\_5\_moe\_a2\_7b\_qlora.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen1_5_moe_a2_7b_qlora.yaml))                                                                                           |
| 2026-01-07 | LLM             | [Devstral-Small-2-24B-Instruct-2512](/model-coverage/large-language-models/mistralai/ministral3-devstral) ([devstral2\_small\_2512\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/devstral/devstral2_small_2512_squad.yaml))                                                   |
| 2025-12-18 | LLM             | [Functiongemma-270m-it](/model-coverage/large-language-models/google/functiongemma) ([functiongemma\_xlam.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gemma/functiongemma_xlam.yaml))                                                                                              |
| 2025-12-16 | LLM             | [Llama-3.1-70B](https://huggingface.co/meta-llama/Llama-3.1-70B) ([llama3\_70b\_pretrain.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_pretrain/llama3_70b_pretrain.yaml))                                                                                                                    |
| 2025-12-05 | VLM             | [Ministral-3-14B-Reasoning-2512](https://huggingface.co/mistralai/Ministral-3-14B-Reasoning-2512) ([ministral3\_14b\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/mistral/ministral3_14b_medpix.yaml))                                                                       |
| 2025-12-05 | VLM             | [Ministral-3-3B-Reasoning-2512](https://huggingface.co/mistralai/Ministral-3-3B-Reasoning-2512) ([ministral3\_3b\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/mistral/ministral3_3b_medpix.yaml))                                                                           |
| 2025-12-05 | VLM             | [Ministral-3-8B-Reasoning-2512](https://huggingface.co/mistralai/Ministral-3-8B-Reasoning-2512) ([ministral3\_8b\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/mistral/ministral3_8b_medpix.yaml))                                                                           |
| 2025-11-24 | LLM             | [GLM-4.5-Air](/model-coverage/large-language-models/thudm/glm-4-moe-glm-4-5-glm-4-7) ([glm\_4.5\_air\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_4.5_air_te_deepep.yaml))                                                                                      |
| 2025-11-24 | VLM             | [Qwen3-VL-30B-A3B-Instruct](https://huggingface.co/Qwen/Qwen3-VL-30B-A3B-Instruct) ([qwen3\_vl\_moe\_30b\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3/qwen3_vl_moe_30b_te_deepep.yaml))                                                                           |
| 2025-11-19 | VLM             | [InternVL3\_5-4B-hf](https://huggingface.co/OpenGVLab/InternVL3_5-4B-hf) ([internvl\_3\_5\_4b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/internvl/internvl_3_5_4b.yaml))                                                                                                          |
| 2025-11-17 | LLM             | [Qwen2.5-32B-Instruct](https://huggingface.co/Qwen/Qwen2.5-32B-Instruct) ([qwen2\_5\_32b\_peft\_benchmark.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen2_5_32b_peft_benchmark.yaml))                                                                                       |
| 2025-11-10 | Omni            | [Qwen3-Omni-30B-A3B-Instruct](/model-coverage/omni/qwen/qwen3-omni) ([qwen3\_omni\_moe\_30b\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3/qwen3_omni_moe_30b_te_deepep.yaml))                                                                                      |
| 2025-10-24 | LLM             | [Qwen3-Next-80B-A3B-Instruct](/model-coverage/large-language-models/qwen/qwen3-next) ([qwen3\_next\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen3_next_te_deepep.yaml))                                                                                        |
| 2025-10-23 | VLM             | [Qwen3-VL-4B-Thinking](https://huggingface.co/Qwen/Qwen3-VL-4B-Thinking) ([qwen3\_vl\_4b\_instruct\_rdr.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3/qwen3_vl_4b_instruct_rdr.yaml))                                                                                          |
| 2025-10-23 | VLM             | [Qwen3-VL-8B-Instruct](/model-coverage/vision-language-models/qwen/qwen3-vl-qwen3-vl-moe) ([qwen3\_vl\_8b\_instruct\_rdr.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen3/qwen3_vl_8b_instruct_rdr.yaml))                                                                         |
| 2025-10-21 | LLM             | [Qwen2.5-7B-Instruct](/model-coverage/large-language-models/qwen/qwen2) ([qwen2\_5\_7b\_instruct\_chat.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen2_5_7b_instruct_chat.yaml))                                                                                            |
| 2025-10-15 | LLM             | [Qwen3-8B](/model-coverage/large-language-models/qwen/qwen3) ([qwen3\_8b\_squad\_spark.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen3_8b_squad_spark.yaml))                                                                                                                |
| 2025-10-05 | LLM             | [Llama-3.3-70B-Instruct](https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct) ([llama\_3\_3\_70b\_instruct\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_3/llama_3_3_70b_instruct_squad.yaml))                                                                    |
| 2025-10-05 | LLM             | [Mixtral-8x7B-v0.1](/model-coverage/large-language-models/mistralai/mixtral) ([mixtral-8x7b-v0-1\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mistral/mixtral-8x7b-v0-1_squad.yaml))                                                                                         |
| 2025-10-01 | LLM             | [Qwen3-30B-A3B](/model-coverage/large-language-models/qwen/qwen3-moe) ([qwen3\_moe\_30b\_te\_deepep.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen3_moe_30b_te_deepep.yaml))                                                                                                |
| 2025-09-29 | LLM             | [DeepSeek-V3](/model-coverage/large-language-models/deepseek-ai/deepseek-v3) ([deepseekv3\_pretrain.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_pretrain/deepseekv3_pretrain.yaml))                                                                                                         |
| 2025-09-29 | LLM             | [Llama-3\_3-Nemotron-Super-49B-v1\_5](https://huggingface.co/nvidia/Llama-3_3-Nemotron-Super-49B-v1_5) ([llama3\_3\_nemotron\_super\_49B\_squad\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/llama3_3_nemotron_super_49B_squad_peft.yaml))                           |
| 2025-09-29 | LLM             | [NVIDIA-Nemotron-Nano-9B-v2](/model-coverage/large-language-models/nvidia/nemotron-h) ([nemotron\_nano\_9b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/nemotron/nemotron_nano_9b_squad.yaml))                                                                               |
| 2025-09-23 | LLM             | [Gpt-oss-120b](/model-coverage/large-language-models/openai/gpt-oss) ([gpt\_oss\_120b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gpt_oss/gpt_oss_120b.yaml))                                                                                                                      |
| 2025-09-23 | LLM             | [Gpt-oss-20b](/model-coverage/large-language-models/openai/gpt-oss) ([gpt\_oss\_20b.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gpt_oss/gpt_oss_20b.yaml))                                                                                                                         |
| 2025-09-08 | LLM             | [Moonlight-16B-A3B](/model-coverage/large-language-models/deepseek-ai/deepseek-v3) ([moonlight\_16b\_te.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/moonlight/moonlight_16b_te.yaml))                                                                                              |
| 2025-09-03 | LLM             | [Gpt2](/model-coverage/large-language-models/openai/gpt-2) ([megatron\_pretrain\_gpt2.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_pretrain/megatron_pretrain_gpt2.yaml))                                                                                                                    |
| 2025-08-27 | LLM             | [OLMo-2-0425-1B-Instruct](/model-coverage/large-language-models/allenai/olmo2) ([olmo\_2\_0425\_1b\_instruct\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/olmo/olmo_2_0425_1b_instruct_squad.yaml))                                                                          |
| 2025-08-27 | LLM             | [Baichuan2-7B-Chat](https://huggingface.co/baichuan-inc/Baichuan2-7B-Chat) ([baichuan\_2\_7b\_mock\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/baichuan/baichuan_2_7b_mock_fp8.yaml))                                                                                         |
| 2025-08-27 | LLM             | [Starcoder2-7b](/model-coverage/large-language-models/bigcode/starcoder2) ([starcoder\_2\_7b\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/starcoder/starcoder_2_7b_hellaswag_fp8.yaml))                                                                             |
| 2025-08-27 | LLM             | [Seed-Coder-8B-Instruct](/model-coverage/large-language-models/bytedance-seed/seed-bytedance) ([seed\_coder\_8b\_instruct\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/seed/seed_coder_8b_instruct_hellaswag_fp8.yaml))                                             |
| 2025-08-27 | LLM             | [Seed-OSS-36B-Instruct](/model-coverage/large-language-models/bytedance-seed/seed-bytedance) ([seed\_oss\_36B\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/seed/seed_oss_36B_hellaswag.yaml))                                                                            |
| 2025-08-27 | LLM             | [C4ai-command-r7b-12-2024](/model-coverage/large-language-models/cohere/command-r) ([cohere\_command\_r\_7b\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/cohere/cohere_command_r_7b_hellaswag_fp8.yaml))                                                            |
| 2025-08-27 | LLM             | [Gemma-2-9b-it](/model-coverage/large-language-models/google/gemma-2) ([gemma\_2\_9b\_it\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gemma/gemma_2_9b_it_hellaswag_fp8.yaml))                                                                                      |
| 2025-08-27 | LLM             | [Gemma-3-270m](https://huggingface.co/google/gemma-3-270m) ([gemma\_3\_270m\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gemma/gemma_3_270m_squad.yaml))                                                                                                                     |
| 2025-08-27 | LLM             | [Gemma-7b](/model-coverage/large-language-models/google/gemma-1) ([gemma\_7b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/gemma/gemma_7b_squad.yaml))                                                                                                                        |
| 2025-08-27 | LLM             | [Granite-3.3-2b-instruct](https://huggingface.co/ibm-granite/granite-3.3-2b-instruct) ([granite\_3\_3\_2b\_instruct\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/granite/granite_3_3_2b_instruct_hellaswag_fp8.yaml))                                               |
| 2025-08-27 | LLM             | [Llama-3.2-3B-Instruct](https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct) ([llama\_3\_2\_3b\_instruct\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_2/llama_3_2_3b_instruct_squad.yaml))                                                                        |
| 2025-08-27 | LLM             | [Phi-2](/model-coverage/large-language-models/microsoft/phi) ([phi\_2\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/phi/phi_2_squad.yaml))                                                                                                                                    |
| 2025-08-27 | LLM             | [Phi-3-mini-4k-instruct](/model-coverage/large-language-models/microsoft/phi-3-phi-4) ([phi\_3\_mini\_it\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/phi/phi_3_mini_it_squad.yaml))                                                                                         |
| 2025-08-27 | LLM             | [Phi-4](https://huggingface.co/microsoft/phi-4) ([phi\_4\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/phi/phi_4_hellaswag_fp8.yaml))                                                                                                                                |
| 2025-08-27 | LLM             | [Mistral-7B-v0.1](/model-coverage/large-language-models/mistralai/mistral) ([mistral\_7b\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mistral/mistral_7b_hellaswag_fp8.yaml))                                                                                       |
| 2025-08-27 | LLM             | [Mistral-Nemo-Base-2407](https://huggingface.co/mistralai/Mistral-Nemo-Base-2407) ([mistral\_nemo\_2407\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mistral/mistral_nemo_2407_hellaswag_fp8.yaml))                                                                 |
| 2025-08-27 | LLM             | [Mixtral-8x7B-Instruct-v0.1](/model-coverage/large-language-models/mistralai/mixtral) ([mixtral\_8x7b\_instruct\_squad\_peft.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/mistral/mixtral_8x7b_instruct_squad_peft.yaml))                                                           |
| 2025-08-27 | LLM             | [Qwen2.5-7B](https://huggingface.co/Qwen/Qwen2.5-7B) ([qwen2\_5\_7b\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen2_5_7b_hellaswag_fp8.yaml))                                                                                                               |
| 2025-08-27 | LLM             | [Qwen3-0.6B](/model-coverage/large-language-models/qwen/qwen3) ([qwen3\_0p6b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwen3_0p6b_hellaswag.yaml))                                                                                                               |
| 2025-08-27 | LLM             | [QwQ-32B](https://huggingface.co/Qwen/QwQ-32B) ([qwq\_32b\_squad.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/qwen/qwq_32b_squad.yaml))                                                                                                                                             |
| 2025-08-27 | LLM             | [Falcon3-7B-Instruct](https://huggingface.co/tiiuae/Falcon3-7B-Instruct) ([falcon3\_7b\_instruct\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/falcon/falcon3_7b_instruct_hellaswag_fp8.yaml))                                                                       |
| 2025-08-27 | LLM             | [Glm-4-9b-chat-hf](/model-coverage/large-language-models/thudm/glm-4) ([glm\_4\_9b\_chat\_hf\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/glm/glm_4_9b_chat_hf_hellaswag_fp8.yaml))                                                                                 |
| 2025-08-23 | LLM             | [Llama-3.1-8B](/model-coverage/large-language-models/meta/llama) ([llama3\_1\_8b\_hellaswag\_fp8.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_1/llama3_1_8b_hellaswag_fp8.yaml))                                                                                             |
| 2025-08-23 | LLM             | [Llama-3.2-1B](/model-coverage/large-language-models/meta/llama) ([llama3\_2\_1b\_hellaswag.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/llm_finetune/llama3_2/llama3_2_1b_hellaswag.yaml))                                                                                                      |
| 2025-08-23 | Omni            | [Phi-4-multimodal-instruct](/model-coverage/omni/microsoft/phi-4-multimodal) ([phi4\_mm\_cv17.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/phi4/phi4_mm_cv17.yaml))                                                                                                                 |
| 2025-08-23 | VLM             | [Gemma-3-4b-it](/model-coverage/vision-language-models/google/gemma-3-vl-gemma-3n) ([gemma3\_vl\_4b\_cord\_v2.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma3/gemma3_vl_4b_cord_v2.yaml))                                                                                       |
| 2025-08-23 | VLM             | [Gemma-3n-e4b-it](https://huggingface.co/google/gemma-3n-e4b-it) ([gemma3n\_vl\_4b\_medpix.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/gemma3n/gemma3n_vl_4b_medpix.yaml))                                                                                                         |
| 2025-08-23 | VLM             | [Qwen2.5-VL-3B-Instruct](/model-coverage/vision-language-models/qwen/qwen2-5-vl) ([qwen2\_5\_vl\_3b\_rdr.yaml](https://github.com/NVIDIA-NeMo/Automodel/blob/main/examples/vlm_finetune/qwen2_5/qwen2_5_vl_3b_rdr.yaml))                                                                                              |