NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (17)
Collection (0)
Model (0)
Resource (0)
Helm Chart (0)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise Supported
14
NVIDIA AI Enterprise
14
NVIDIA Developer Program
14
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
17
NIM Container GPUs
(1)
NIM Container GPUs
NIM Container GPUs
H100 80GB HBM3
28
L40S
24
A100 SXM4 80GB
22
B200
20
H200
17
H100 NVL
15
A10G
11
A100 PG509 200
10
GB200
7
GH200 120GB
7
GH200 480GB
7
DGX Spark
6
H200 NVL
6
RTX PRO 6000 Blackwell Server Edition
6
GH200 144G HBM3e
5
H100 PCIe
4
L4
4
RTX PRO 6000 Blackwell Workstation Edition
3
B300 SXM6 AC
2
L40
2
RTX PRO 6000 Blackwell Max Q Workstation Edition
2
GB300
1
Use Case
(0)
Use Case
Use Case
Natural Language Processing
1
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
Industry
(0)
Industry
Industry
Solution
(0)
Solution
Solution
Conversational AI
1
Publisher
(0)
Publisher
Publisher
Nvidia
12
Google
1
Jet artifacts registry
1
Moonshotai
1
Stepfun ai
1
Stockmark
1
Policy
(0)
Policy
Policy
Displaying 17 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
H200
NIM Container GPUs: H200
Clear Filters
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
mixtral-8x7b-instruct-v0-1
Please add descriptionNVIDIA NIM for GPU accelerated Mixtral-8x7B-Instruct-v0.1 inference through OpenAI compatible APIs
A100 PG509 200
A100 SXM4 80GB
+7
A10G
B200
H100 80GB HBM3
H100 NVL
H200
L40
L40S
Container
14mo
Updated
06/04/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Mistral-7B-Instruct-v0.3
This container houses the Mistral-7B-Instruct-v0.3, a large language model fine-tuned for instruction-based tasks. An improved version of Mistral-7B-v0.3, it is designed to be easily fine-tuned to achieve compelling performance.
A100 PG509 200
A100 SXM4 80GB
+13
A10G
B200
GB200
GH200 120GB
GH200 144G HBM3e
GH200 480GB
H100 80GB HBM3
H100 NVL
H200
H200 NVL
L40
L40S
RTX PRO 6000 Blackwell Server Edition
Container
11mo
Updated
09/04/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama-3.2-1B-Instruct
This container houses the Llama-3.2-1B-Instruct, a compact 1B parameter model from Meta AI fine-tuned for dialogue. It's designed to be highly efficient and accessible, making it adept at question answering, summarization, and other instructions.
A100 PG509 200
A100 SXM4 80GB
+12
A10G
B200
GB200
GH200 120GB
GH200 144G HBM3e
GH200 480GB
H100 80GB HBM3
H100 NVL
H200
H200 NVL
L40S
RTX PRO 6000 Blackwell Server Edition
Container
11mo
Updated
08/27/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama-3.2-3B-Instruct
NVIDIA NIM for GPU accelerated Llama-3.2-3B-Instruct inference through OpenAI compatible APIs
A100 PG509 200
A100 SXM4 80GB
+8
A10G
B200
GH200 120GB
GH200 480GB
H100 80GB HBM3
H100 NVL
H200
L40S
Container
12mo
Updated
07/28/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama-3.3-Nemotron-Super-49B-v1
NVIDIA NIM for GPU accelerated Llama-3.3-Nemotron-Super-49B-v1 inference through OpenAI compatible APIs
A100 PG509 200
A100 SXM4 80GB
+8
A10G
B200
GH200 120GB
GH200 480GB
H100 80GB HBM3
H100 NVL
H200
L40S
Container
12mo
Updated
07/16/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama-3.1-Nemotron-Nano-8B-v1
NVIDIA NIM for GPU accelerated Llama-3.1-Nemotron-Nano-8B-v1 inference through OpenAI compatible APIs
A100 SXM4 80GB
A10G
+4
H100 80GB HBM3
H100 NVL
H200
L40S
Container
14mo
Updated
06/02/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
NVIDIA-Nemotron-Nano-12B-v2-VL
This container houses the Nemotron Nano12B v2 VL model, which is an advanced autoregressive Visual Language Model (VLM) designed for document transcription from images and videos, outputting text in a reading order.
B200
GB200
+12
GH200 120GB
GH200 144G HBM3e
GH200 480GB
H100 80GB HBM3
H100 NVL
H100 PCIe
H200
H200 NVL
L40S
RTX PRO 6000 Blackwell Max Q Workstation Edition
RTX PRO 6000 Blackwell Server Edition
RTX PRO 6000 Blackwell Workstation Edition
Container
6mo
Updated
01/27/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Cosmos Reason-2-8B
Given a text prompt and an input video, think and generate the answer with respect to the input text prompt and video.
B200
B300 SXM6 AC
+14
DGX Spark
GB200
GH200 120GB
GH200 144G HBM3e
GH200 480GB
H100 80GB HBM3
H100 NVL
H100 PCIe
H200
H200 NVL
L40S
RTX PRO 6000 Blackwell Max Q Workstation Edition
RTX PRO 6000 Blackwell Server Edition
RTX PRO 6000 Blackwell Workstation Edition
Container
3mo
Updated
04/13/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
Stockmark
Stockmark-2-100B-Instruct
NVIDIA NIM for GPU accelerated Stockmark-2-100B-Instruct inference through OpenAI compatible APIs
A100 PG509 200
A100 SXM4 80GB
+11
B200
Conversational AI
GB200
GH200 144G HBM3e
H100 80GB HBM3
H100 NVL
H200
H200 NVL
L40S
Natural Language Processing
RTX PRO 6000 Blackwell Server Edition
Container
10mo
Updated
09/25/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama 3.1 Nemotron Nano VL 8B v1
This container houses the Llama-3.1-Nemotron-Nano-VL-8B-V1 model, a leading document intelligence vision language model (VLMs) that enables the ability to query and summarize images from the physical or virtual world.
H100 80GB HBM3
H100 NVL
+3
H100 PCIe
H200
L40S
Container
12mo
Updated
07/30/2025 UTC
NVIDIA
minimax-m2-5
The MiniMax-M2.5 NIM Container is a deployable inference container for serving MiniMax-M2.5, a third-party text generation model optimized for complex agentic tasks including software engineering, tool use, search.
B200
DGX Spark
+3
GB300
H100 80GB HBM3
H200
Container
1mo
Updated
06/25/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
JET Artifacts Registry
qwen3-next-80b-a3b-instruct
Qwen3-Next-80B-A3B is the first installment in the Qwen3-Next series and features the following key enchancements: Hybrid Attention, High-Sparsity Mixture-of-Experts (MoE), Stability Optimizations and Multi-Token Prediction (MTP)
B200
H100 80GB HBM3
+1
H200
Container
8mo
Updated
12/03/2025 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Qwen3-Next-80B-A3B-Thinking
NVIDIA NIM for GPU accelerated Qwen3-Next-80B-A3B inference through OpenAI compatible APIs
B200
H100 80GB HBM3
+1
H200
Container
5mo
Updated
02/11/2026 UTC
Google
Gemma 4 31B IT
Gemma 4 31B IT model which, is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output.
B200
H200
+1
L40S
Container
1mo
Updated
06/25/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
MoonshotAI
Kimi K2.6
The Kimi-K2.6 Certified NIM Container is a deployable inference container for serving Kimi-K2.6, a third-party multimodal Mixture-of-Experts model capable of processing text, image, and video inputs for text generation.
B200
H100 80GB HBM3
+1
H200
Container
2mo
Updated
06/01/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Cosmos3 Reasoner
Given multimodal inputs including text, images, and video, generates coherent text responses for multimodal understanding, action reasoning, and Physical AI applications.
B200
B300 SXM6 AC
+11
GB200
GH200 120GB
GH200 480GB
H100 80GB HBM3
H100 NVL
H100 PCIe
H200
H200 NVL
L40S
RTX PRO 6000 Blackwell Server Edition
RTX PRO 6000 Blackwell Workstation Edition
Container
2mo
Updated
05/30/2026 UTC
Stepfun-AI
Step 3.7 Flash
The Step-3.7-Flash NIM is a Downloadable NIM container for deploying Step-3.7-Flash, a StepFun vision-language model built on Step 3.5 Flash with additional vision capability for native multimodal, agentic, and coding-related use cases.
B200
H100 80GB HBM3
+1
H200
Container
2mo
Updated
05/28/2026 UTC
24
Select item
24
48
96
192
24
48
96
192
1-17 of 17 items
1
1
π