NVIDIA
NVIDIA
NeMo NIM Proxy
Helm Chart
NVIDIA
NVIDIA
NeMo NIM Proxy

Proxy service for NIM microservices

Sign in to access all content for this Helm ChartSigning in will also allow download accessSign In

NeMo NIM Proxy Microservice Helm Chart

NeMo NIM Proxy microservice provides a unified access point for all NVIDIA NIM (NVIDIA Inference Microservice) deployments within your Kubernetes cluster through a single OpenAI-compatible API.

You can use the NIM Proxy microservice to interact with multiple NIM microservices through a unified endpoint proxy API. With this microservice, you can retrieve the deployed NIM microservices and make inference requests to the chat/completions and completions APIs.

Alternative: Platform Deployment

You can install NeMo NIM Proxy as part of the NeMo microservices platform by using the NeMo Microservices Helm Chart (chart | documentation).

Resources

Container | Helm Installation Guide | User Guide

Note: Use, distribution or deployment of this microservice in production requires an NVIDIA AI Enterprise License.

Governing Terms

The software and materials are governed by the NVIDIA Software License Agreement and the Product-Specific Terms for NVIDIA AI Products.

Publisher
NVIDIA
NVIDIA
Latest Version25.6.0
UpdatedJune 11, 2025 UTC
Compressed Size31.71 KB

NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.