NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (13)
Collection (4)
Model (10)
Resource (2)
Helm Chart (5)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA Developer Program
3
NVIDIA AI Enterprise Supported
2
NVIDIA AI Enterprise
2
NVIDIA NeMo Microservices
1
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
2
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
Use Case
(1)
Use Case
Use Case
Automatic Speech Recognition
119
Natural Language Processing
104
Natural Language Understanding
64
Drug Discovery
61
Language Modeling
37
Simulation and Modeling
37
Question Answering
34
Video Analytics
34
Object Detection
33
Speech to Text
33
Text to Speech
26
High Performance Computing
25
Video enhancement
24
Recommendation
23
Application Development
22
Translation
22
Synthetic Data Generation
15
Forecasting
13
Image Segmentation
11
Audio Synthesis
10
Speech enhancement
10
Annotation
8
GPU Enablement with Kubernetes
8
Genome Sequencing
8
Graph Neural Networks
8
Action Recognition
5
Image Synthesis
4
Reinforcement Learning
4
Named Entity Recognition
3
Facial Landmark Estimation
2
Body Pose Classification
1
Body Pose Estimation
1
Emotion Classification
1
Eye Gaze Estimation
1
Gesture Classification
1
Heart Rate Estimation
1
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
NeMo
6
Metropolis Microservices
3
Metropolis
2
Riva
2
CUDA
1
CUDA Toolkit
1
Deep Learning Examples
1
DeepStream
1
PyTorch
1
RAPIDS
1
TensorFlow
1
TensorRT
1
Industry
(0)
Industry
Industry
Retail
4
Smart Cities / Spaces
4
Automotive / Transportation
2
Robotics
2
HPC / Supercomputing
1
Solution
(0)
Solution
Solution
AI
22
NVIDIA AI
10
Conversational AI
7
Vision AI
7
Computer Vision
6
Inference
5
DL
4
High Performance Computing
2
ML
2
Developer Tools
1
Kubernetes Infrastructure
1
Recommender Systems
1
Tools and Management
1
Publisher
(0)
Publisher
Publisher
Nvidia
29
Nvidia corporation
1
Policy
(0)
Policy
Policy
Displaying 34 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
Question Answering
Use Case: Question Answering
Clear Filters
NVIDIA
vLLM
vLLM is a fast and easy-to-use library for LLM inference and serving. The NVIDIA vLLM NGC Container is optimized for GPU acceleration, and contains a validated set of libraries that enable and optimize GPU performance.
AI
Conversational AI
+10
DL
High Performance Computing
HPC / Supercomputing
Inference
ML
Natural Language Processing
Natural Language Understanding
NVIDIA AI
Question Answering
Translation
Container
1w
Updated
07/27/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
Llama-3.2-11B-Vision-Instruct
The Llama 3.2 Vision instruction-tuned models are optimized for visual recognition, image reasoning, captioning, and answering general questions about an image.
Automotive / Transportation
Computer Vision
+3
Image Segmentation
Question Answering
Vision AI
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
VSS Engine
Build a Video Search and Summarization Agent Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A
AI
Question Answering
+2
Video Analytics
Vision AI
Container
6mo
Updated
01/28/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
meta-llama-2-7b-chat
NVIDIA NIM for GPU accelerated Llama 2 7B inference through OpenAI compatible APIs
Automotive / Transportation
Language Modeling
+1
Question Answering
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
SGLang
SGLang is a fast serving framework for large language models and vision language models. The NVIDIA SGLang NGC Container is optimized for GPU acceleration, and contains a validated set of libraries that enable and optimize GPU performance.
AI
Conversational AI
+9
DL
High Performance Computing
Inference
ML
Natural Language Processing
Natural Language Understanding
NVIDIA AI
Question Answering
Translation
Container
1w
Updated
07/27/2026 UTC
NVIDIA
AI-Q Research Assistant Backend
The NVIDIA AI-Q Research Assistant Blueprint gives developers a foundational starting point for building a deep research assistant that can run on-premise. The backend container provides the RESTful API service.
AI
NVIDIA AI
+1
Question Answering
Container
8mo
Updated
11/06/2025 UTC
NVIDIA
VSS Engine Base Container
Base Container for building VSS engine from source
AI
Question Answering
+2
Video Analytics
Vision AI
Container
6mo
Updated
01/28/2026 UTC
NVIDIA
AI-Q Research Assistant Frontend
The NVIDIA AI-Q Research Assistant Blueprint gives developers a foundational starting point for building a deep research assistant that can run on-premise. The frontend container provides a demo web UI.
AI
NVIDIA AI
+1
Question Answering
Container
8mo
Updated
11/06/2025 UTC
NVIDIA NeMo Microservices
+1
NVIDIA Developer Program
NVIDIA
NVIDIA NVIngest Microservice
Helm Chart for NeMo Retriever NVIngest Microservice
AI
NeMo
+1
Question Answering
Helm Chart
4mo
Updated
03/17/2026 UTC
NVIDIA
AI Blueprint for Video Search and Summarization
Blueprint for the Video Search and Summarization Agent
AI
Computer Vision
+6
Metropolis Microservices
Question Answering
Retail
Smart Cities / Spaces
Video Analytics
Vision AI
Helm Chart
6mo
Updated
01/28/2026 UTC
NVIDIA
NVIDIA K8s Developer LLM Operator
The NVIDIA K8s Developer LLM Operator is an open source and easy to deploy Kubernetes Operator to self-host Generative AI workflows.
AI
Conversational AI
+6
Developer Tools
Inference
Kubernetes Infrastructure
NVIDIA AI
Question Answering
Tools and Management
Helm Chart
2y
Updated
07/08/2024 UTC
NVIDIA
AI-Q Research Assistant File Loader
The NVIDIA AI-Q Research Assistant Blueprint gives developers a foundational starting point for building a deep research assistant that can run on-premise. This container includes a utility to create two default datasets used by the demo web UI.
AI
NVIDIA AI
+1
Question Answering
Container
8mo
Updated
11/06/2025 UTC
NVIDIA
VLM Inference Service (Jetson)
AI Inference Service for using VLM (visual language model) on streaming video for greater contextual understanding and natural language interaction
AI
Computer Vision
+10
Language Modeling
Metropolis
Metropolis Microservices
Question Answering
Restaurant / Quick-Service
Retail
Robotics
Smart Cities / Spaces
Video Analytics
Vision AI
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
aiq-agent
NVIDIA AI-Q Intelligence Agent — an enterprise-grade backend agent built on the NVIDIA NeMo Agent Toolkit, providing quick cited answers and in-depth report-style research with modular multi-agent workflows.
AI
Natural Language Processing
+2
NeMo
Question Answering
Container
2mo
Updated
05/20/2026 UTC
NVIDIA
aiq-frontend
NVIDIA AI-Q Blueprint frontend — a modern web UI built with Next.js, React, and NVIDIA KUI Foundations, providing an accessible interface for the AI-Q Blueprint backend with optional OAuth authentication.
AI
Natural Language Processing
+3
NeMo
NVIDIA AI
Question Answering
Container
1d
Updated
08/06/2026 UTC
NVIDIA
NeMo Retriever Library (NRL) Service
NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice.
AI
Conversational AI
+2
NeMo
Question Answering
Container
1mo
Updated
06/30/2026 UTC
NVIDIA
NVIDIA AI-Q Research Assistant Blueprint Helm Chart
Helm chart to deploy the NVIDIA AI-Q Research Assistant Blueprint which gives developers a starting point for building a deep research assistant that can run on-premise, allowing anyone to create detailed research reports using on-premise data.
AI
NVIDIA AI
+1
Question Answering
Helm Chart
6mo
Updated
01/16/2026 UTC
NVIDIA Corporation
AI-Q Blueprint
AI-Q Blueprint Helm Chart
AI
Natural Language Processing
+2
Natural Language Understanding
Question Answering
Helm Chart
1mo
Updated
06/09/2026 UTC
NVIDIA
Jetson Platform Services Reference Workflow & Resources
Jetson Platform Services Reference Workflow & Resources
AI
Computer Vision
+15
Config file
DeepStream
Inference
Metropolis
Metropolis Microservices
Object Detection
Question Answering
Restaurant / Quick-Service
Retail
Robotics
Script
Smart Cities / Spaces
Tutorial
Video Analytics
Vision AI
Resource
18mo
Updated
01/15/2025 UTC
NVIDIA
QA squadv2.0 Bertbase
Question answering model with BERT base encoder finetuned on SQuADv2.0
Question Answering
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
QA squadv1.1 Bertbase
Question answering model with BERT base encoder finetuned on SQuADv1.1
Natural Language Processing
Question Answering
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
QA squadv1.1 Megatronuncased
Uncased question answering model with Megatron encoder finetuned on SQuADv1.1
Natural Language Processing
Question Answering
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
QA squadv2.0 Bertlarge
Question answering model with BERT large encoder finetuned on SQuADv2.0
Question Answering
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
QA squadv1.1 Megatroncased
Cased question answering model with Megatron encoder finetuned on SQuADv1.1
Natural Language Processing
Question Answering
Model
>3y
Updated
04/04/2023 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 34 items
1
1
2
2
π