NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (27)
Collection (0)
Model (0)
Resource (0)
Helm Chart (0)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise Supported
1
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
27
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
B200
3
H200
3
H100 80GB HBM3
2
DGX Spark
1
GB300
1
L40S
1
Use Case
(0)
Use Case
Use Case
Natural Language Processing
3
Natural Language Understanding
3
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
NeMo
1
Industry
(0)
Industry
Industry
Solution
(0)
Solution
Solution
AI
3
NVIDIA AI
2
Publisher
(0)
Publisher
Publisher
Nvidia
14
Google
3
Qwen
3
Deepseek ai
1
Moonshot ai
1
Stepfun ai
1
Thinking machines lab
1
Xiaomi
1
Z.ai
1
Zai org
1
Policy
(0)
Policy
Policy
Displaying 27 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
Experimental
label: Experimental
Clear Filters
NVIDIA
minimax-m2-5
The MiniMax-M2.5 NIM Container is a deployable inference container for serving MiniMax-M2.5, a third-party text generation model optimized for complex agentic tasks including software engineering, tool use, search.
B200
DGX Spark
+3
GB300
H100 80GB HBM3
H200
Container
1mo
Updated
06/25/2026 UTC
NVIDIA
gliner-pii
This container houses GLiNER PII, which detects and classifies a broad range of Personally Identifiable Information (PII) and Protected Health Information (PHI) in structured and unstructured text.
NVIDIA AI
Container
1mo
Updated
06/25/2026 UTC
NVIDIA
Nemotron-3-Super-120B-A12B (Experimental)
Nemotron-3-Super-120B-A12B is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.
Container
1mo
Updated
06/25/2026 UTC
Google
Gemma 4 26B A4B IT
Gemma 4 26B A4B IT is a Google multimodal instruction-tuned model packaged as an NVIDIA NIM container for deployment through NVIDIA NGC as a Downloadable NIM.
Container
2mo
Updated
06/02/2026 UTC
Qwen
Qwen3.5-35B-A3B
Qwen3.5-35B-A3B is a multimodal vision-language Mixture-of-Experts model designed for native multimodal agent applications, supporting text, image, and video inputs.
Container
1mo
Updated
06/25/2026 UTC
Google
Gemma 4 31B IT
Gemma 4 31B IT model which, is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output.
B200
H200
+1
L40S
Container
1mo
Updated
06/25/2026 UTC
NVIDIA
step-35-flash
Step 3.5 Flash is a sparse Mixture-of-Experts (MoE) large language model developed by StepFun, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency
NVIDIA AI
Container
5mo
Updated
02/20/2026 UTC
Qwen
Qwen3.6-35B-A3B
The Qwen3.6-35B-A3B NIM Container is a deployable inference container for serving Qwen3.6-35B-A3B, a third-party multimodal Mixture of Experts model capable of processing text, image, and video inputs for text generation.
Container
4d
Updated
08/06/2026 UTC
Moonshot AI
kimi-k2.5-Turbo
This turbo container houses the Kimi K2.5 model which is an open-source, native multimodal agentic model built through continual pretraining on approximately 15 trillion mixed visual and text tokens atop Kimi-K2-Base.
Container
1mo
Updated
06/25/2026 UTC
zai-org
glm-5.2
This container houses GLM-5.2, a flagship long-context large language model for agentic engineering and advanced reasoning, packaged as an NVIDIA NIM for staging.
Container
1mo
Updated
07/10/2026 UTC
NVIDIA
Nemotron-3-Nano-Omni-30B-A3B-Reasoning
Nemotron Nano V3 Omni is a multi-modal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
Container
1mo
Updated
06/22/2026 UTC
Z.Ai
GLM-5
GLM-5 is a next-generation large language model targeting complex systems engineering and long-horizon agentic tasks.
Container
5mo
Updated
02/13/2026 UTC
NVIDIA
Nemotron-Content-Safety-Reasoning-4B (Experimental)
Nemotron Content Safety Reasoning 4B is a Large Language Model (LLM) classifier designed to function as a dynamic and adaptable guardrail for content safety and dialogue moderation (topic-following).
AI
Natural Language Processing
+2
Natural Language Understanding
NeMo
Container
3mo
Updated
04/24/2026 UTC
NVIDIA
nvidia-ising-calibration-1
The NVIDIA Ising Calibration 1 NIM houses the NVIDIA-Ising-Calibration-1-35B-A3B-BF16 model, which is a purpose-built Mixture-of-Experts vision-language model (MoE VLM) built on Qwen3.5-35B-A3B,
Container
3mo
Updated
04/15/2026 UTC
NVIDIA
GLM-5.1
This container houses GLM-5.1, which is a next-generation flagship model for agentic engineering with significantly stronger coding capabilities than its predecessor GLM-5. The model achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5
Container
1mo
Updated
06/11/2026 UTC
Xiaomi
Mimo-V2-Flash (Experimental)
This container houses the model MiMo-V2-Flash.
AI
Natural Language Processing
+1
Natural Language Understanding
Container
3mo
Updated
04/29/2026 UTC
NVIDIA
nemotron-3.5-content-safety
The Nemotron 3.5 Content Safety NIM container packages NVIDIA's small language model (SLM) that uses Google's Gemma-3-4B-it as the base and is fine-tuned by NVIDIA on multimodal, multilingual, and reasoning-oriented content-safety datasets.
Container
2mo
Updated
06/01/2026 UTC
Google
DiffusionGemma 4 26B A4B IT
The DiffusionGemma-4-26B-A4B-IT model is an open-weights multimodal generative model developed by Google DeepMind that processes text, image, and video inputs to produce text output via discrete diffusion.
Container
1mo
Updated
06/25/2026 UTC
Qwen
qwen3.6-27b
The Qwen3.6-27B NIM Container is a deployable inference container for serving Qwen3.6-27B, a third-party multimodal dense model capable of processing text, image, and video inputs for text generation.
Container
3mo
Updated
04/23/2026 UTC
NVIDIA
Nemotron-3 Content Safety VLM
This NIM container houses the Nemotron 3 Content Safety model which, is a small language model (SLM) that uses Google's Gemma-3-4B-it as the base and is fine-tuned by NVIDIA on multimodal and multilingual content-safety related datasets.
Container
1mo
Updated
06/25/2026 UTC
Deepseek AI
DeepSeek-V4-Pro
The DeepSeek-V4-Pro Container is a deployable inference container for serving DeepSeek-V4-Pro, a third-party sparse Mixture-of-Experts language model for reasoning, coding, and agentic tasks.
Container
3mo
Updated
04/25/2026 UTC
NVIDIA
Nemotron-3-Super-120B-A12B (Turbo)
Nemotron-3-Super-120B-A12B is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.
Container
1mo
Updated
06/25/2026 UTC
Stepfun-AI
Step 3.7 Flash
The Step-3.7-Flash NIM is a Downloadable NIM container for deploying Step-3.7-Flash, a StepFun vision-language model built on Step 3.5 Flash with additional vision capability for native multimodal, agentic, and coding-related use cases.
B200
H100 80GB HBM3
+1
H200
Container
2mo
Updated
05/28/2026 UTC
Thinking Machines Lab
Inkling
This NIM container houses Inkling model, which is a 66-layer decoder-only transformer with a sparse Mixture-of-Experts (MoE) feed-forward backbone, featuring 975B total parameters and 41B active parameters
Container
3w
Updated
07/16/2026 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 27 items
1
1
2
2
π