Enterprise RAG Deployment Guide#
NVIDIA Enterprise Reference Architecture
- Abstract
- Introduction
- Enterprise Reference Architecture Overview
- Enterprise RAG on Enterprise RA 285
- Deploy Enterprise RAG on Enterprise RA
- Deploy the RAG Ingestion Pipeline
- Configure MIG
- Configure Run:ai
- MIG: NeMo Retriever (NV-Ingest) with MIG Profiles
- Run:ai: NeMo Retriever (NV-Ingest) with GPU Fractions
- Install NV-Ingest client
- Dataset - Download Enterprise dataset
- Deploy Vector database (Milvus)
- Deploy Milvus (Distributed) with GPU index and CPU search using MIG or Run:ai
- Deploy RAG Retrieval pipeline
- Deploy RAG Blueprint
- Install AI-Perf benchmarking tool
- Run AI-Perf Benchmarking tool
- Summary - RAG Deployment Best Practices
- RAG Ingestion Deployment
- References
- Appendix
- MIG - NeMo Retriever (NV-Ingest) MIG custom Helm values file
- Run:ai 1 - NeMo Retriever (NV-Ingest) custom Helm values file
- Run:ai 2 - NeMo Retriever (NV-Ingest) NIM manifest deployment
- Milvus distributed Helm custom values file
- NIM Cache for Retrieval NIM
- NIM Service for Retrieval NIM microservices
- RAG Blueprint Helm custom values
Notices