NIM LLM and VLM Documentation# About NIM LLM and VLM Overview NIM Offerings Enterprise-Grade Inference Software Stack Release Notes Get Started About Get Started Prerequisites Configuration Installation Quickstart Advanced Deployment Model Profiles and Selection Model Download Model-Free NIM Kubernetes Deployment Cloud Service Provider (CSP) Deployment Air-Gap Deployment Multi-Node Deployment vGPU Deployment Advanced Use Cases Fine-Tuning with LoRA Speculative Decoding Payload Capture Tool Calling and MCP Integration Custom Parsers and Chat Templates Custom Logits Processing Prompt Embeddings Image, Audio, and Video Input AI Assistant Integrations Use Claude Code with NIM Use Codex CLI with NIM Reference Benchmarking Architecture Environment Variables API Reference CLI Reference Advanced Configuration Logging and Observability Model Signature Verification 1.x Migration Guide Support Matrix for NIMs Archived Versions Troubleshooting GPU Memory (OOM) Errors Request Timeouts and Responses CUDA Driver Initialization OpenShift SCC Admission Failures Resources Support and FAQ Related Products Legal and Acknowledgements