NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (65)
Collection (8)
Model (4)
Resource (25)
Helm Chart (10)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise Supported
20
NVIDIA AI Enterprise
13
NVIDIA Mission Control
11
NVIDIA Developer Program
4
NVIDIA AI Enterprise IGX
2
Cosmos Transfer1 7B
1
NVIDIA Maxine Early Access
1
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
3
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
Use Case
(0)
Use Case
Use Case
Natural Language Understanding
13
Natural Language Processing
12
Object Detection
10
Automatic Speech Recognition
7
Question Answering
5
Recommendation
4
Text to Speech
3
Translation
3
Video Analytics
3
Application Development
2
Drug Discovery
2
Image Segmentation
2
Action Recognition
1
Body Pose Classification
1
Body Pose Estimation
1
Emotion Classification
1
Eye Gaze Estimation
1
Facial Landmark Estimation
1
GPU Enablement with Kubernetes
1
Gesture Classification
1
Graph Neural Networks
1
Heart Rate Estimation
1
High Performance Computing
1
Named Entity Recognition
1
Reinforcement Learning
1
Simulation and Modeling
1
Speech to Text
1
Synthetic Data Generation
1
Video enhancement
1
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
Triton Inference Server
16
Morpheus
7
PyTorch
6
TensorFlow
5
DeepStream
4
Merlin
4
NeMo
4
TensorRT
3
CUDA
2
CUDA Toolkit
2
Holoscan
2
Metropolis
2
Metropolis Microservices
2
RAPIDS
2
Clara
1
Clara AGX
1
HPC
1
Maxine
1
Omniverse
1
PyTorch Geometric
1
Runs on RTX
1
TAO Toolkit
1
Industry
(0)
Industry
Industry
Academia / Higher Education
8
Automotive / Transportation
7
Cloud Services
7
Financial Services
7
Public Sector
7
Retail
7
HPC / Supercomputing
4
Healthcare
4
Robotics
4
Smart Cities / Spaces
3
Life Sciences
2
Media & Entertainment
2
Agriculture
1
Energy
1
Manufacturing
1
Solution
(1)
Solution
Solution
AI
397
DL
270
NVIDIA AI
258
Infrastructure Software
238
Conversational AI
176
Kubernetes Infrastructure
169
Inference
112
ML
111
Computer Vision
103
High Performance Computing
55
Tools and Management
49
Vision AI
46
Developer Tools
40
Application Development
16
GPU Accelerated Libraries
16
Rendering
14
Genomics
12
Graphics and Simulation
11
Physics and Dynamics Simulation
11
Data Analytics
9
Languages and APIs
8
Application Streaming
4
HPC Benchmarks
3
Recommender Systems
3
Data Center Simulation Platform
2
OpenACC Programming Model
1
Scientific Visualization
1
Publisher
(0)
Publisher
Publisher
Nvidia
100
Meta
2
Gcp
1
Mit
1
Preferred networks
1
Policy
(0)
Policy
Policy
Government ready
Labeled versions meet security requirements for FedRAMP High or equivalent use cases
4
Displaying 112 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
Inference
Solution: Inference
Clear Filters
NVIDIA
Triton Inference Server
Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.
Automatic Speech Recognition
Automotive / Transportation
+6
DL
Inference
Infrastructure Software
NVIDIA AI
Object Detection
Triton Inference Server
Container
1w
Updated
07/29/2026 UTC
NVIDIA
TensorRT
NVIDIA TensorRT is a C++ library that facilitates high-performance inference on NVIDIA graphics processing units (GPUs). TensorRT takes a trained network and produces a highly optimized runtime engine that performs inference for that network.
Automotive / Transportation
DL
+2
Inference
NVIDIA AI
Container
1w
Updated
07/27/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
Nvidia
NVIDIA NIM for GenMol
GenMol is a masked diffusion model trained on molecular SAFE representations for fragment-based molecule generation, which can serve as a generalist model for various drug discovery tasks.
AI
Drug Discovery
+5
Healthcare
Inference
Life Sciences
NVIDIA AI
PyTorch
Container
2mo
Updated
05/07/2026 UTC
NVIDIA
vLLM
vLLM is a fast and easy-to-use library for LLM inference and serving. The NVIDIA vLLM NGC Container is optimized for GPU acceleration, and contains a validated set of libraries that enable and optimize GPU performance.
AI
Conversational AI
+10
DL
High Performance Computing
HPC / Supercomputing
Inference
ML
Natural Language Processing
Natural Language Understanding
NVIDIA AI
Question Answering
Translation
Container
1w
Updated
07/27/2026 UTC
NVIDIA
NVIDIA NIM Operator
An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
AI
Inference
+2
NeMo
NVIDIA AI
Container
2mo
Updated
05/20/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
MIT
DiffDock
Diffdock predicts the 3D structure of the interaction between a molecule and a protein.
AI
CUDA
+10
CUDA Toolkit
Drug Discovery
Healthcare
Inference
Life Sciences
NVIDIA AI
PyTorch
PyTorch Geometric
RAPIDS
Triton Inference Server
Container
4w
Updated
07/10/2026 UTC
NVIDIA
Dynamo vLLM Runtime
The Dynamo vLLM runtime image is a containerized build of Dynamo + vLLM which serves as the base runtime environment for vLLM based inference with Dynamo's distributed inference framework.
AI
DL
+5
High Performance Computing
Inference
Infrastructure Software
ML
NVIDIA AI
Container
1d
Updated
08/05/2026 UTC
NVIDIA
Dynamo Tensorrt-LLM Runtime
The Dynamo TensorRT-LLM runtime image is a containerized build of Dynamo + TensorRT-LLM which serves as the base runtime environment for tensorrt-llm based inference with Dynamo's distributed inference framework.
AI
DL
+5
High Performance Computing
Inference
Infrastructure Software
ML
NVIDIA AI
Container
1d
Updated
08/05/2026 UTC
NVIDIA
CUDA GL
CUDA is a parallel computing platform and programming model that enables dramatic increases in computing performance by harnessing the power of the NVIDIA GPUs. These images extend the CUDA images to include OpenGL support through libglvnd.
DL
HPC
+2
Inference
Infrastructure Software
Container
8mo
Updated
11/06/2025 UTC
NVIDIA
Dynamo Platform
A comprehensive Helm chart for deploying the NVIDIA Dynamo operator and its dependencies
AI
DL
+4
Inference
Kubernetes Infrastructure
ML
NVIDIA AI
Helm Chart
1d
Updated
08/05/2026 UTC
NVIDIA
NVIDIA NIM Operator
Helm chart for NIM Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
AI
Inference
+3
Kubernetes Infrastructure
NeMo
NVIDIA AI
Helm Chart
2mo
Updated
05/20/2026 UTC
NVIDIA
Dynamo SGLang Runtime
The Dynamo SGLang runtime image is a containerized build of Dynamo + SGLang which serves as the base runtime environment for sglang based inference with Dynamo's distributed inference framework.
AI
DL
+5
High Performance Computing
Inference
Infrastructure Software
ML
NVIDIA AI
Container
1d
Updated
08/05/2026 UTC
NVIDIA
Dynamo kubernetes-operator
kubernetes-operator is a container that runs as part of the Dynamo cloud platform. Dynamo cloud is a kubernetes platform for deploying and managing inference services. This container manages the lifecycle of Dynamo inference deployments in kubernetes.
AI
DL
+5
High Performance Computing
Inference
Infrastructure Software
ML
NVIDIA AI
Container
1d
Updated
08/05/2026 UTC
NVIDIA
Kaldi
Kaldi is an open-source software framework for speech processing.
Automatic Speech Recognition
Conversational AI
+6
DL
Inference
ML
Natural Language Processing
Natural Language Understanding
Retail
Container
8mo
Updated
11/06/2025 UTC
NVIDIA
Riva Skills Quick Start
Scripts and utilities for getting started with Riva Speech Skills
AI
Automatic Speech Recognition
+4
Conversational AI
Inference
Natural Language Understanding
Text to Speech
Resource
16mo
Updated
04/04/2025 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB October 2024 (PB 24h2)
Triton Inference Server Production Branch October 2024 (PB 24h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
8mo
Updated
11/06/2025 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB October 2025 (PB 25h2)
Triton Inference Server PB October 2025 (PB 25h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
2w
Updated
07/23/2026 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB March 2025 (PB 25h1)
Triton Inference Server PB May 2025 (PB 25h1) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
7mo
Updated
12/17/2025 UTC
NVIDIA
NVCF ClusterAgent
NVCF ClusterAgent (NVCA) is a self-installable micro-service that enables a compute backend, DGX Clouds or other NVIDIA Compute Backends to be used as an NVCF Target to host NVIDIA Cloud Functions’ instances.
Inference
Kubernetes Infrastructure
+1
NVIDIA AI
Container
2w
Updated
07/23/2026 UTC
NVIDIA
Merlin PyTorch
The Merlin PyTorch container allows users to do preprocessing and feature engineering with NVTabular, and then train a deep-learning based recommender system model with PyTorch, and serve the trained model on Triton Inference Server.
Automotive / Transportation
DL
+7
Inference
Merlin
ML
NVIDIA AI
PyTorch
Recommendation
Triton Inference Server
Container
22mo
Updated
09/25/2024 UTC
NVIDIA
Morpheus
NVIDIA Morpheus is an open AI application framework for cybersecurity developers.
Academia / Higher Education
AI
+13
Cloud Services
Data Analytics
Developer Tools
DL
Financial Services
GPU Accelerated Libraries
Inference
Languages and APIs
ML
Morpheus
Natural Language Processing
NVIDIA AI
Public Sector
Container
6mo
Updated
01/21/2026 UTC
NVIDIA
Dynamo CRDs
A Helm chart that manages Custom Resource Definitions (CRDs) for the NVIDIA Dynamo ecosystem in Kubernetes
AI
DL
+4
Inference
Kubernetes Infrastructure
ML
NVIDIA AI
Helm Chart
5mo
Updated
03/04/2026 UTC
NVIDIA
NVCA Webhook Server
NVCF ClusterAgent (NVCA) is a self-installable micro-service that enables a compute backend to host NVIDIA Cloud Functions’ instances. NVCA Webhook Server is used by NVCA to enforce restrictions for NVCF Mini Service.
AI
Inference
+1
Kubernetes Infrastructure
Container
8mo
Updated
11/10/2025 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server Long-Term Support Branch 2 (LTSB 2)
Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.
Inference
Triton Inference Server
Container
1w
Updated
07/29/2026 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 112 items
1
1
2
2
3
3
4
4
5
5
π