NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
Llama-3.1-Nemotron-70B-Instruct
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (289)
Collection (16)
Model (4)
Resource (5)
Helm Chart (48)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise Supported
166
NVIDIA AI Enterprise
143
NVIDIA Developer Program
121
NIMs for China Region
3
NVIDIA Earth-2 Inference
2
NVIDIA NeMo Microservices
2
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
362
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
H100 80GB HBM3
25
L40S
21
B200
20
A100 SXM4 80GB
19
H200
17
H100 NVL
15
A10G
11
A100 PG509 200
10
GB200
7
GH200 120GB
7
GH200 480GB
7
H200 NVL
6
RTX PRO 6000 Blackwell Server Edition
6
GH200 144G HBM3e
5
H100 PCIe
4
L4
4
DGX Spark
3
RTX PRO 6000 Blackwell Workstation Edition
3
B300 SXM6 AC
2
L40
2
RTX PRO 6000 Blackwell Max Q Workstation Edition
2
GB300
1
Use Case
(0)
Use Case
Use Case
Speech to Text
19
Video enhancement
18
Translation
14
Automatic Speech Recognition
13
Drug Discovery
12
Speech enhancement
9
Natural Language Processing
8
Natural Language Understanding
7
Forecasting
4
Simulation and Modeling
4
Text to Speech
4
High Performance Computing
3
Language Modeling
3
Synthetic Data Generation
3
Genome Sequencing
2
Graph Neural Networks
2
Image Segmentation
2
Image Synthesis
2
Question Answering
2
Video Analytics
2
Annotation
1
Application Development
1
Audio Synthesis
1
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
Riva
22
Holoscan
14
Maxine
11
Triton Inference Server
11
TensorRT
7
PyTorch
6
Clara
5
PhysicsNeMo
5
NeMo
3
CUDA
2
Monai
2
CUDA Toolkit
1
HPC
1
Metropolis
1
Morpheus
1
PyTorch Geometric
1
RAPIDS
1
Industry
(0)
Industry
Industry
Automotive / Transportation
42
Media & Entertainment
14
Healthcare
11
Consumer Internet
9
Life Sciences
9
HPC / Supercomputing
3
Robotics
3
Telecommunications
2
Academia / Higher Education
1
Aerospace
1
Cloud Services
1
Manufacturing
1
Retail
1
Smart Cities / Spaces
1
Solution
(0)
Solution
Solution
Conversational AI
30
NVIDIA AI
29
AI
14
Computer Vision
12
Vision AI
10
DL
9
Genomics
4
Inference
4
High Performance Computing
3
Physics and Dynamics Simulation
3
ML
1
Time Series
1
Publisher
(0)
Publisher
Publisher
Nvidia
291
Qwen
6
Google
5
Meta
5
Mistral ai
5
Mit
3
Alibaba
2
Arc
2
Deepseek ai
2
Deepseekai
2
Hive
2
Ipd
2
Moonshot ai
2
Openai
2
Colabfold
1
Jet artifacts registry
1
Llama 3.1 nemoguard 8b content safety
1
Microsoft
1
Moonshotai
1
Opengpt x
1
Sarvamai
1
Stability ai
1
Stepfun ai
1
Stockmark
1
Thinking machines lab
1
Wan ai
1
Xiaomi
1
Z.ai
1
Zai org
1
Policy
(0)
Policy
Policy
Government ready
Labeled versions meet security requirements for FedRAMP High or equivalent use cases
16
Displaying 362 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
nimmcro_nvidia_nim
label: nimmcro_nvidia_nim
Clear Filters
NVIDIA
Llama-3.1-70b-instruct PB October 2024 (PB 24h2)
Llama 3.1 70B-Instruct NIM Production Branch October 2024 (PB 24h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
Automotive / Transportation
Container
9mo
Updated
11/06/2025 UTC
NVIDIA AI Enterprise
—
Snowflake Arctic Embed Large Embedding
NVIDIA NIM for GPU accelerated Snowflake Arctic Embed Large Embedding inference
Automotive / Transportation
Container
2y
Updated
07/25/2024 UTC
NVIDIA
Gemma-2-2B-IT
NVIDIA NIM for GPU accelerated Gemma-2-2B-IT inference through OpenAI compatible APIs
Automotive / Transportation
Container
19mo
Updated
01/08/2025 UTC
NVIDIA
meta-llama-2-70b-chat
NVIDIA NIM for GPU accelerated Llama 2 70B inference through OpenAI compatible APIs
Automotive / Transportation
Synthetic Data Generation
+1
Translation
Container
2y
Updated
08/20/2024 UTC
NVIDIA
Llama-3-Swallow-70B-Instruct-v0.1
NVIDIA NIM for GPU accelerated Llama-3-Swalow-70B-Instruct-v0.1 inference through OpenAI compatible APIs
Automotive / Transportation
Container
23mo
Updated
08/27/2024 UTC
NVIDIA
Llama-3.1-Swallow-8B-Instruct-v0.1
NVIDIA NIM for GPU accelerated Llama 3.1 Swallow 8B inference through OpenAI compatible APIs
Automotive / Transportation
Container
20mo
Updated
12/20/2024 UTC
NVIDIA
Mistral-Nemo-12B-Instruct
NVIDIA NIM for GPU accelerated Mistral-NeMo-12B-Instruct inference through OpenAI compatible APIs
Container
17mo
Updated
03/12/2025 UTC
Hive
deepfake-image-detection
Hive’s Deepfake Image Detection model analyzes images and returns a confidence score on how likely the image contains a deepfake.
Container
18mo
Updated
01/23/2025 UTC
NVIDIA
Llama-3.1-8b-instruct PB October 2024 (PB 24h2)
NVIDIA NIM for GPU accelerated Llama 3.1 8B inference through OpenAI compatible APIs
Container
9mo
Updated
11/17/2025 UTC
NVIDIA AI Enterprise
Nvidia
NVIDIA NIM for GenMol
GenMol is a masked diffusion model trained on molecular SAFE representations for fragment-based molecule generation, which can serve as a generalist model for various drug discovery tasks.
AI
Drug Discovery
+5
Healthcare
Inference
Life Sciences
NVIDIA AI
PyTorch
Container
3mo
Updated
05/07/2026 UTC
NVIDIA
Audio2Face-3D
NVIDIA NIM for GPU accelerated Audio2Face-3D inference through gRPC APIs.
Container
5mo
Updated
03/11/2026 UTC
NVIDIA
Llama-3.1-70b-instruct
This container houses the Llama-3.1-70B-Instruct, which is a multilingual large language model from the Meta Llama 3.1 collection of pretrained and instruction-tuned generative models.
Container
2d
Updated
08/20/2026 UTC
NVIDIA
CodeLlama-70B-Instruct
NVIDIA NIM for GPU accelerated CodeLlama-70B inference through OpenAI compatible APIs
Application Development
Automotive / Transportation
+1
Simulation and Modeling
Container
22mo
Updated
10/07/2024 UTC
NVIDIA
NeMo Retriever PaddleOCR
PaddleOCR is an ultra lightweight Optical Character Recognition (OCR) system by Baidu. PaddleOCR supports a variety of cutting-edge algorithms related to OCR.
A100 PG509 200
A100 SXM4 80GB
+6
A10G
B200
H100 80GB HBM3
H100 NVL
L4
L40S
Container
11mo
Updated
09/05/2025 UTC
NVIDIA
Parakeet 0.6b CTC en-US NIM
Parakeet 0.6b CTC en-US NIM delivers accurate English speech-to-text transcription and enables easy-to-use optimized ASR inference for large scale deployments.
AI
Conversational AI
+2
Riva
Speech to Text
Container
8mo
Updated
12/04/2025 UTC
NVIDIA
Llama-3.1-8B-Instruct
NVIDIA NIM for GPU accelerated Llama-3.1-8B-Instruct inference through OpenAI compatible APIs
Container
2d
Updated
08/19/2026 UTC
NVIDIA
Deepseek-R1-Distill-Qwen-32B
NVIDIA NIM for GPU accelerated DeepSeek-R1-Distill-Qwen-32B inference through OpenAI compatible APIs
Container
17mo
Updated
03/20/2025 UTC
NVIDIA
Deepseek-R1-Distill-Llama-8B
NVIDIA NIM for GPU accelerated DeepSeek-R1-Distill-Llama-8B inference through OpenAI compatible APIs
Container
17mo
Updated
02/26/2025 UTC
NVIDIA
NV-CLIP
NV-CLIP NIM microservice for multimodal embeddings model for image and text
AI
Annotation
+7
Automotive / Transportation
Computer Vision
Metropolis
Retail
Robotics
Smart Cities / Spaces
Video Analytics
Container
17mo
Updated
03/03/2025 UTC
NVIDIA
Riva NMT NIM
Riva NMT NIM provide easy access to state-of-the-art neural machine translation (NMT) models, capable of translating text from one language to another with exceptional accuracy.
Automotive / Transportation
Conversational AI
+2
Riva
Translation
Container
18mo
Updated
02/17/2025 UTC
NVIDIA
Llama-3.2-11B-Vision-Instruct
The Llama 3.2 Vision instruction-tuned models are optimized for visual recognition, image reasoning, captioning, and answering general questions about an image.
Automotive / Transportation
Computer Vision
+3
Image Segmentation
Question Answering
Vision AI
Container
19mo
Updated
01/17/2025 UTC
MIT
DiffDock
Diffdock predicts the 3D structure of the interaction between a molecule and a protein.
AI
CUDA
+10
CUDA Toolkit
Drug Discovery
Healthcare
Inference
Life Sciences
NVIDIA AI
PyTorch
PyTorch Geometric
RAPIDS
Triton Inference Server
Container
1mo
Updated
07/10/2026 UTC
NVIDIA
Llama 3.1 NemoGuard 8B Topic Control
NVIDIA NIM for GPU accelerated Llama 3.1 NemoGuard 8B Topic Control inference through OpenAI compatible APIs
Container
12mo
Updated
08/07/2025 UTC
NVIDIA
Llama-3-Taiwan-70B-Instruct
NVIDIA NIM for GPU accelerated Llama-3-Taiwan-70B-Instruct inference through OpenAI compatible APIs
Automotive / Transportation
Container
23mo
Updated
08/27/2024 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 362 items
1
1
2
2
3
3
4
4
5
5
...
16
16
π