country_code
Skip to main content
Ctrl+K
NVIDIA NIM for LLM and VLM - Home NVIDIA NIM for LLM and VLM - Home

NVIDIA NIM for LLM and VLM

  • Documentation Home
NVIDIA NIM for LLM and VLM - Home NVIDIA NIM for LLM and VLM - Home

NVIDIA NIM for LLM and VLM

  • Documentation Home

Table of Contents

About NIM LLM and VLM

  • Overview
  • NIM Offerings
  • Enterprise-Grade Inference Software Stack
  • Release Notes

Get Started

  • About Get Started
  • Prerequisites
  • Configuration
  • Installation
  • Quickstart
  • Advanced
    • Model-Specific Deployment Considerations
    • Get Started with DeepSeek-V4-Pro-0813

Deployment

  • Model Profiles and Selection
  • Model Download
  • Model-Free NIM
  • Kubernetes Deployment
    • Helm and Kubernetes
    • KServe
    • OpenShift
    • Run:ai
    • NIM Operator Deployment
  • Cloud Service Provider (CSP) Deployment
    • Google Cloud
    • AWS
    • Azure
    • Oracle
  • Air-Gap Deployment
  • Multi-Node Deployment
  • vGPU Deployment

Advanced Use Cases

  • Fine-Tuning with LoRA
  • Speculative Decoding
  • Payload Capture
  • Tool Calling and MCP Integration
  • Custom Parsers and Chat Templates
  • Custom Logits Processing
  • Prompt Embeddings
  • Image, Audio, and Video Input

AI Assistant Integrations

  • Use Claude Code with NIM
  • Use Codex CLI with NIM

Reference

  • Benchmarking
  • Architecture
  • Environment Variables
  • API Reference
  • CLI Reference
  • Advanced Configuration
  • Logging and Observability
  • Model Signature Verification
  • 1.x Migration Guide
  • Support Matrix for NIMs
  • Archived Versions

Troubleshooting

  • GPU Memory (OOM) Errors
  • Request Timeouts and Responses
  • CUDA Driver Initialization
  • OpenShift SCC Admission Failures

Resources

  • Support and FAQ
  • Related Products
  • Legal
  • Advanced
Is this page helpful?

Advanced#

Some NIMs require more involved setup than the standard Get Started path. Use the following guides for those NIMs.

Note

If you are new to NIM LLM and VLM, start with the Get Started documentation before using these guides.

Model-Specific Deployment Considerations

Model-Specific Deployment Considerations

how-to

Model-Specific Deployment Considerations
DeepSeek-V4-Pro-0813

Get Started with DeepSeek-V4-Pro-0813

how-to

Get Started with DeepSeek-V4-Pro-0813

previous

Quickstart

next

Model-Specific Deployment Considerations

NVIDIA NVIDIA
Privacy Policy | Your Privacy Choices | Terms of Service | Accessibility | Corporate Policies | Product Security | Contact

Copyright © 2024-2026, NVIDIA Corporation.

Last updated on Sep 24, 2026.