NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (17)
Collection (7)
Model (2)
Resource (8)
Helm Chart (0)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise
15
NVIDIA AI Enterprise Supported
10
NVIDIA Developer Program
10
NVIDIA AI Enterprise IGX
1
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
10
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
Use Case
(0)
Use Case
Use Case
Automatic Speech Recognition
8
Speech to Text
7
Recommendation
4
Natural Language Understanding
3
Object Detection
3
Drug Discovery
2
Natural Language Processing
2
Graph Neural Networks
1
High Performance Computing
1
Image Segmentation
1
Simulation and Modeling
1
Video Analytics
1
Video enhancement
1
NVIDIA Platform
(1)
NVIDIA Platform
NVIDIA Platform
NeMo
149
Omniverse
106
Riva
85
Deep Learning Examples
78
PyTorch
59
DOCA
51
Maxine
47
DeepStream
43
Metropolis
42
Clara
39
TAO Toolkit
36
Runs on RTX
35
Triton Inference Server
34
Holoscan
32
Isaac
26
TensorRT
25
Aerial
24
HPC
19
TensorFlow
19
PhysicsNeMo
17
CUDA
15
Network Operator
12
Metropolis Microservices
11
Morpheus
11
RAPIDS
11
Merlin
8
GPU Operator
5
CUDA Toolkit
4
Container Toolkit
4
Clara AGX
3
Clara Parabricks
3
Deep Learning Institute
3
Monai
3
cuOpt
3
DCGM
2
HPC SDK
2
PyTorch Geometric
2
Deep Graph Library
1
GPU Driver
1
JAX
1
Industry
(0)
Industry
Industry
Automotive / Transportation
4
Robotics
2
HPC / Supercomputing
1
Healthcare
1
Life Sciences
1
Media & Entertainment
1
Solution
(0)
Solution
Solution
Inference
16
NVIDIA AI
11
AI
10
Conversational AI
9
DL
8
Computer Vision
5
ML
3
Vision AI
2
High Performance Computing
1
Infrastructure Software
1
Kubernetes Infrastructure
1
Physics and Dynamics Simulation
1
Recommender Systems
1
Publisher
(0)
Publisher
Publisher
Nvidia
26
Mit
1
Policy
(0)
Policy
Policy
Government ready
Labeled versions meet security requirements for FedRAMP High or equivalent use cases
2
Displaying 34 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
Triton Inference Server
NVIDIA Platform: Triton Inference Server
Clear Filters
NVIDIA
Triton Inference Server
Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.
Automatic Speech Recognition
Automotive / Transportation
+6
DL
Inference
Infrastructure Software
NVIDIA AI
Object Detection
Triton Inference Server
Container
1w
Updated
07/29/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
MIT
DiffDock
Diffdock predicts the 3D structure of the interaction between a molecule and a protein.
AI
CUDA
+10
CUDA Toolkit
Drug Discovery
Healthcare
Inference
Life Sciences
NVIDIA AI
PyTorch
PyTorch Geometric
RAPIDS
Triton Inference Server
Container
4w
Updated
07/10/2026 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB October 2024 (PB 24h2)
Triton Inference Server Production Branch October 2024 (PB 24h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
9mo
Updated
11/06/2025 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB October 2025 (PB 25h2)
Triton Inference Server PB October 2025 (PB 25h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
2w
Updated
07/23/2026 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server PB March 2025 (PB 25h1)
Triton Inference Server PB May 2025 (PB 25h1) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
7mo
Updated
12/17/2025 UTC
NVIDIA
Merlin PyTorch
The Merlin PyTorch container allows users to do preprocessing and feature engineering with NVTabular, and then train a deep-learning based recommender system model with PyTorch, and serve the trained model on Triton Inference Server.
Automotive / Transportation
DL
+7
Inference
Merlin
ML
NVIDIA AI
PyTorch
Recommendation
Triton Inference Server
Container
22mo
Updated
09/25/2024 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server Long-Term Support Branch 2 (LTSB 2)
Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.
Inference
Triton Inference Server
Container
1w
Updated
07/29/2026 UTC
NVIDIA
Merlin TensorFlow
The Merlin TensorFlow container allows users to do preprocessing and feature engineering with NVTabular, and then train a deep-learning based recommender system model with TensorFlow, and serve the trained model on Triton Inference Server.
Automotive / Transportation
DL
+6
Inference
Merlin
NVIDIA AI
Recommender Systems
TensorFlow
Triton Inference Server
Container
22mo
Updated
09/25/2024 UTC
NVIDIA AI Enterprise IGX
NVIDIA
Tritonserver LTSB2 IGX
Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.Please add description
Inference
Triton Inference Server
Container
1w
Updated
07/29/2026 UTC
NVIDIA
Merlin HugeCTR
The Merlin HugeCTR container enables you to perform data preprocessing, feature engineering, train models with HugeCTR, and then serve the trained model with Triton Inference Server.
Automotive / Transportation
DL
+6
Inference
Merlin
ML
NVIDIA AI
Recommendation
Triton Inference Server
Container
2y
Updated
07/17/2024 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
relighting
AI4M Relighting is an AI-powered video relighting that dynamically re-illuminates a person with virtual studio lighting using HDR environment maps. Supports adjustable lighting direction, intensity, specular highlights, and background compositing.
Computer Vision
Media & Entertainment
+3
TensorRT
Triton Inference Server
Video enhancement
Container
3mo
Updated
04/17/2026 UTC
NVIDIA
Merlin PyTorch Inference
This container allows users to deploy NVTabular workflows and PyTorch models to Triton Inference server for production.
Merlin
PyTorch
+2
Recommendation
Triton Inference Server
Container
>4y
Updated
05/12/2022 UTC
NVIDIA AI Enterprise
NVIDIA
Triton Inference Server Production Branch 6
Triton Inference Server Production Branch 6 offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
AI
DL
+3
Inference
NVIDIA AI
Triton Inference Server
Container
1w
Updated
07/31/2026 UTC
NVIDIA
Nvidia VSS CV Event Detector
Nvidia Sample CV Event Detector Microservice for detecting events for VSS Event Reviewer workflow.
AI
Computer Vision
+8
DeepStream
Metropolis
Metropolis Microservices
Object Detection
TensorRT
Triton Inference Server
Video Analytics
Vision AI
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
Merlin Tensorflow Inference
This container allows users to deploy NVTabular workflows and TensorFlow models to Triton Inference server for production.
Inference
Merlin
+3
Recommendation
TensorFlow
Triton Inference Server
Container
>4y
Updated
05/12/2022 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
ALCHEMI Batched Molecular Dynamics
Atomistic Molecular Dynamics NIM with machine learning interatomic potentials.
Drug Discovery
Graph Neural Networks
+6
High Performance Computing
HPC / Supercomputing
Physics and Dynamics Simulation
PyTorch
Simulation and Modeling
Triton Inference Server
Container
4mo
Updated
03/19/2026 UTC
NVIDIA
Fine-Tune and Optimize BERT
Jupyter Notebooks for BERT Pre-training, Fine-Tuning and Inference profiling and optimization via TensorFlow, AMP, XLA, DLProf, TF-TRT and Triton.
Conversational AI
Natural Language Processing
+2
Natural Language Understanding
Triton Inference Server
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
BERT on Google Cloud AI Platform
Fine-tune a pre-trained BERT model with the SQuAD dataset, optimize for inference using TensorRT and deploy with Triton Inference Server on Google Cloud AI Platform using Custom Containers
AI
Conversational AI
+4
Inference
Natural Language Processing
Natural Language Understanding
Triton Inference Server
Resource
>3y
Updated
04/04/2023 UTC
NVIDIA
Traffic Cam Analyzer on A100 MIG
Build an example Traffic Cam Footage Analyzer Using the NVIDIA DeepStream SDK and Triton Inference Server to identify and classify vehicles (sedan, truck, SUV, etc.) from live traffic camera streams
Computer Vision
DeepStream
+2
Inference
Triton Inference Server
Resource
>3y
Updated
04/04/2023 UTC
NVIDIA
PeopleSemSeg AMR
People semantic segmentation network, finetuned on robotics AMR dataset, optimized for Issac Perceptor & Nvblox.
AI
Computer Vision
+7
Image Segmentation
NVIDIA AI
Robotics
TAO Toolkit
TensorRT
Triton Inference Server
Vision AI
Model
20mo
Updated
12/03/2024 UTC
NVIDIA
OpenShift BERT Example
This notebook demonstrates how to optimize a fine-tuned BERT TF checkpoint to TensorRT and then how to deploy it for inference using Triton inference server on OpenShift cluster.
Inference
Kubernetes Infrastructure
+2
TensorFlow
Triton Inference Server
Resource
>3y
Updated
04/04/2023 UTC
NVIDIA
PeopleNet AMR
People bounding box detection network, finetuned on robotics AMR dataset, optimized for multi-camera RealSense setup. Used in nvblox multi-camera optimization.
AI
Computer Vision
+8
Inference
ML
NVIDIA AI
Object Detection
Robotics
TAO Toolkit
TensorRT
Triton Inference Server
Model
20mo
Updated
12/03/2024 UTC
NVIDIA
BERT QA on Azure ML with Triton Demo
Demo notebook to deploy BERT QA model on Azure ML with Triton Inference Server
Natural Language Understanding
Triton Inference Server
Resource
>3y
Updated
04/04/2023 UTC
NVIDIA
triton-sdk-2.52.0-rhel8-x86-compat-beta.zip
RHEL Triton x86 SDK Beta asset
Triton Inference Server
Resource
19mo
Updated
12/13/2024 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 34 items
1
1
2
2
π