SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA Enterprise
NVIDIA Enterprise
49
34
11
10
3
2
1
1
1
1
1
NVIDIA NIM
NVIDIA NIM
13
NIM Container GPUs
NIM Container GPUs
Use Case
Use Case
59
37
23
23
22
21
17
16
9
7
6
6
6
6
5
5
5
4
4
4
4
2
2
2
1
1
1
1
1
1
1
1
1
1
1
1
NVIDIA Platform
NVIDIA Platform
48
30
26
19
17
15
15
13
10
10
9
7
7
6
5
5
4
4
4
3
3
2
2
1
1
1
1
1
1
1
1
1
Industry
Industry
58
25
24
23
18
16
14
11
9
9
8
6
6
5
5
5
4
4
2
1
Solution
Solution
392
266
243
222
173
155
108
107
103
47
47
46
40
16
16
14
12
12
10
9
8
4
3
3
2
1
1
Publisher
Publisher
292
12
10
8
6
5
3
2
1
1
1
1
1
1
1
1
Policy
Policy
9
Displaying 392 results
Docker containers distributed as part of the TAO Toolkit package
Container
NVIDIA
NVIDIA
PyTorch
PyTorch is a GPU accelerated tensor computational framework. Functionality can be extended with common Python libraries such as NumPy and SciPy. Automatic differentiation is done with a tape-based system at the functional and neural network layer levels.
Container
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs.
Container
DeepStream SDK delivers a complete streaming analytics toolkit for AI based video and image understanding and multi-sensor processing. This container is for NVIDIA Enterprise GPUs.
Container
NVIDIA Developer Program
GenMol is a masked diffusion model trained on molecular SAFE representations for fragment-based molecule generation, which can serve as a generalist model for various drug discovery tasks.
Container
NVIDIA
NVIDIA
vLLM
vLLM is a fast and easy-to-use library for LLM inference and serving. The NVIDIA vLLM NGC Container is optimized for GPU acceleration, and contains a validated set of libraries that enable and optimize GPU performance.
Container
An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
Container
NVIDIA Developer Program
Parakeet 0.6b CTC en-US NIM delivers accurate English speech-to-text transcription and enables easy-to-use optimized ASR inference for large scale deployments.
Container
CUDA is a parallel computing platform and programming model that enhances computing performance using NVIDIA GPUs. CUDA Deep Learning integrates networking and GPU-accelerated libraries like cuDNN, cuTensor, NCCL, HPC-x, and the CUDA Toolkit.
Container
NVIDIA Developer Program
NVIDIA
NVIDIA
NV-CLIP
NV-CLIP NIM microservice for multimodal embeddings model for image and text
Container
NVIDIA Developer Program
Diffdock predicts the 3D structure of the interaction between a molecule and a protein.
Container
This is the RAG server container used as part of the NVIDIA AI Blueprint for RAG and used to orchestrate the end to end RAG pipeline.
Container
NVIDIA Developer Program
NVIDIA
NVIDIA
MolMIM
MolMIM is a transformer-based model developed by NVIDIA for controlled small molecule generation.
Container
The Dynamo vLLM runtime image is a containerized build of Dynamo + vLLM which serves as the base runtime environment for vLLM based inference with Dynamo's distributed inference framework.
Container
The Dynamo TensorRT-LLM runtime image is a containerized build of Dynamo + TensorRT-LLM which serves as the base runtime environment for tensorrt-llm based inference with Dynamo's distributed inference framework.
Container
NVIDIA
NVIDIA
JAX
JAX is a framework for high-performance numerical computing and machine learning research. It includes Numpy-like APIs, automatic differentiation, XLA acceleration and simple primitives for scaling across GPUs and supports an ecosystem of libraries.
Container
This is the Ingestor server container used as part of the NVIDIA RAG Blueprint and used to orchestrate the end to end Ingestion.
Container
NVIDIA
NVIDIA
PyG
PyG (PyTorch Geometric) is a library built upon PyTorch to easily write and train Graph Neural Networks (GNNs) for a wide range of applications related to structured data.
Container
Helm chart for NIM Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
Helm Chart
A comprehensive Helm chart for deploying the NVIDIA Dynamo operator and its dependencies
Helm Chart
Build a Video Search and Summarization Agent Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A
Container
The Dynamo SGLang runtime image is a containerized build of Dynamo + SGLang which serves as the base runtime environment for sglang based inference with Dynamo's distributed inference framework.
Container
kubernetes-operator is a container that runs as part of the Dynamo cloud platform. Dynamo cloud is a kubernetes platform for deploying and managing inference services. This container manages the lifecycle of Dynamo inference deployments in kubernetes.
Container
BioNeMo Framework for running training and inference on large scale bio-based models.
Container
...

NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.