NGC Catalog
Explore
Search
Support
API Catalog
Forum
Search
Containers
DeepSeek-R1
NVIDIA Developer Program
+1
Llama-3.1-Nemotron-70B-Instruct
NVIDIA Developer Program
+1
PyTorch
Collections
Omniverse Kit (FB)
NVIDIA AI Enterprise
+2
DeepStream SDK
Omniverse Kit App Streaming
NVIDIA AI Enterprise
+2
Models
StyleGAN3 pretrained models
PeopleNet
TrafficCamNet
Resources
Riva Skills Quick Start
Helm Charts
GPU Operator
NVIDIA NIM Operator
Welcome Guest
Setup
Terms of Use
Theme
Use System Settings
Light
Dark
Sign In / Sign Up
Search
Search thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
Search
Container (7)
Collection (0)
Model (25)
Resource (5)
Helm Chart (0)
NVIDIA AI Enterprise
(0)
NVIDIA AI Enterprise
NVIDIA AI Enterprise
NVIDIA AI Enterprise Supported
4
NVIDIA ACE Early Access
3
NVIDIA AI Enterprise
2
Mistral Nemotron 4B Instruct
1
NVIDIA Developer Program
1
NVIDIA NIM
(0)
NVIDIA NIM
NVIDIA NIM
Accelerate custom generative AI app deployment using pre-built containers with optimized AI models.
NVIDIA NIM
3
NIM Container GPUs
(0)
NIM Container GPUs
NIM Container GPUs
Use Case
(1)
Use Case
Use Case
Automatic Speech Recognition
119
Natural Language Processing
104
Natural Language Understanding
64
Drug Discovery
61
Language Modeling
37
Simulation and Modeling
37
Question Answering
34
Video Analytics
34
Object Detection
33
Speech to Text
33
Text to Speech
26
High Performance Computing
25
Video enhancement
24
Recommendation
23
Application Development
22
Translation
22
Synthetic Data Generation
15
Forecasting
13
Image Segmentation
11
Audio Synthesis
10
Speech enhancement
10
Annotation
8
GPU Enablement with Kubernetes
8
Genome Sequencing
8
Graph Neural Networks
8
Action Recognition
5
Image Synthesis
4
Reinforcement Learning
4
Named Entity Recognition
3
Facial Landmark Estimation
2
Body Pose Classification
1
Body Pose Estimation
1
Emotion Classification
1
Eye Gaze Estimation
1
Gesture Classification
1
Heart Rate Estimation
1
NVIDIA Platform
(0)
NVIDIA Platform
NVIDIA Platform
Riva
13
NeMo
8
PyTorch
4
Runs on RTX
3
Deep Learning Examples
2
Metropolis
1
Metropolis Microservices
1
Industry
(0)
Industry
Industry
Gaming
5
Healthcare
4
Hardware / Semiconductor
3
Automotive / Transportation
2
HPC / Supercomputing
1
Retail
1
Robotics
1
Smart Cities / Spaces
1
Solution
(0)
Solution
Solution
AI
16
NVIDIA AI
16
DL
9
Conversational AI
5
Computer Vision
1
High Performance Computing
1
ML
1
Vision AI
1
Publisher
(0)
Publisher
Publisher
Nvidia
30
Servicenow hugging face nvidia
2
University of florida health
2
Microsoft
1
Policy
(0)
Policy
Policy
Government ready
Labeled versions meet security requirements for FedRAMP High or equivalent use cases
1
Displaying 37 results
Sort: Most Popular
Sort: Most Popular
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Sort: Relevance
Sort: Most Popular
Sort: Last Updated
Sort: Alphabetical (A-Z)
Sort: Alphabetical (Z-A)
Search
Language Modeling
Use Case: Language Modeling
Clear Filters
NVIDIA
PyTorch
PyTorch is a GPU accelerated tensor computational framework. Functionality can be extended with common Python libraries such as NumPy and SciPy. Automatic differentiation is done with a tape-based system at the functional and neural network layer levels.
AI
Automotive / Transportation
+9
DL
High Performance Computing
HPC / Supercomputing
Language Modeling
ML
Natural Language Processing
Natural Language Understanding
NVIDIA AI
PyTorch
Container
1w
Updated
07/27/2026 UTC
NVIDIA
Llama-3.3-nemotron-super-49b-v1.5
This container houses the Llama-3.3-Nemotron-Super-49B-v1.5, which is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct
Language Modeling
Natural Language Processing
+1
Natural Language Understanding
Container
2d
Updated
08/05/2026 UTC
NVIDIA Developer Program
+1
NVIDIA AI Enterprise
NVIDIA
meta-llama-2-7b-chat
NVIDIA NIM for GPU accelerated Llama 2 7B inference through OpenAI compatible APIs
Automotive / Transportation
Language Modeling
+1
Question Answering
Container
9mo
Updated
11/06/2025 UTC
NVIDIA
nemo-rl
NVIDIA NeMo™ RL accelerates reinforcement learning post-training with high-performance GPU backends, offering scalable GRPO, DPO, SFT, and distillation for multimodal models from single-node experiments to enterprise-scale clusters.
AI
Deep Learning Examples
+4
Language Modeling
NeMo
NVIDIA AI
Reinforcement Learning
Container
1w
Updated
07/28/2026 UTC
NVIDIA
nemo-automodel
NVIDIA NeMo™ AutoModel accelerates LLM and VLM training and fine‑tuning with PyTorch DTensor‑native SPMD, day‑0 Hugging Face support, and optimized parallelism from single‑ to multi‑node scale.
AI
Deep Learning Examples
+4
Language Modeling
NeMo
NVIDIA AI
PyTorch
Container
1mo
Updated
07/02/2026 UTC
NVIDIA AI Enterprise
NVIDIA
Llama-3.3-nemotron-super-49b-v1.5-PB-25h2
This container houses the **Llama-3.3-Nemotron-Super-49B-v1.5 PB 25h2**, which is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct
Language Modeling
Natural Language Processing
+1
Natural Language Understanding
Container
2w
Updated
07/23/2026 UTC
Microsoft
Phi-4
Phi-4 is a state-of-the-art open model built upon a blend of synthetic datasets, data from filtered public domain websites, and acquired academic books and Q&A datasets. Source: https://huggingface.co/microsoft/phi-4
Language Modeling
Model
12mo
Updated
07/17/2025 UTC
NVIDIA
VLM Inference Service (Jetson)
AI Inference Service for using VLM (visual language model) on streaming video for greater contextual understanding and natural language interaction
AI
Computer Vision
+10
Language Modeling
Metropolis
Metropolis Microservices
Question Answering
Restaurant / Quick-Service
Retail
Robotics
Smart Cities / Spaces
Video Analytics
Vision AI
Container
9mo
Updated
11/06/2025 UTC
University of Florida Health
GatorTron-OG
GatorTron-OG is a Megatron BERT model trained on pre-trained on de-identified clinical notes from the University of Florida Health System.
AI
Conversational AI
+5
DL
Healthcare
Language Modeling
Natural Language Processing
Natural Language Understanding
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
BioMegatron345mUncased
Megatron pretrained on uncased biomedical dataset PubMed with 345 million parameters.
AI
Conversational AI
+7
DL
Healthcare
Language Modeling
Natural Language Processing
Natural Language Understanding
NeMo
PyTorch
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
ASR Language Modeling Transformer Large LibriSpeech
Transformer-Large language model for English ASR, Trained on LibriSpeech text corpus with NeMo
Language Modeling
NeMo
Model
>3y
Updated
04/04/2023 UTC
University of Florida Health
GatorTron-S
GatorTron-S is a Megatron BERT model trained on pre-trained on synthetic clinical discharge summaries generated by SynGatorTron 5B NLG, a Megatron GPT-3 model trained on de-identified clinical free text at the University of Florida health system.
AI
Conversational AI
+9
DL
Google Cloud
Healthcare
Language Modeling
Natural Language Processing
Natural Language Understanding
Quick Deploy
Vertex AI
Vertex AI Workbench
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
Riva ASR Arabic LM
Base Arabic 4-gram LM
Arabic
Language Modeling
+1
NVIDIA AI
Model
>3y
Updated
04/18/2023 UTC
NVIDIA
Riva ASR German LM
Base German 4-gram LM
German
Language Modeling
+1
NVIDIA AI
Model
>2y
Updated
09/07/2023 UTC
NVIDIA
RIVA ASR English en-GB LM
Base English en-GB n-gram LM
English
Language Modeling
+2
NVIDIA AI
Riva
Model
13mo
Updated
07/03/2025 UTC
—
Llama2 (DGXC Benchmarking)
This recipe contains information and scripts to produce performance results for the Llama 2 training workload.
AI
Config file
+7
DL
English
Hardware / Semiconductor
Language Modeling
Natural Language Processing
NeMo
Script
Resource
19mo
Updated
12/23/2024 UTC
NVIDIA
Riva ASR Japanese LM
Base Japanese 3-gram LM
Japanese
Language Modeling
+2
NVIDIA AI
Riva
Model
>3y
Updated
06/26/2023 UTC
NVIDIA
Riva ASR Mandarin LM
Base Mandarin 4-gram LM
Language Modeling
Mandarin
+1
NVIDIA AI
Model
13mo
Updated
07/03/2025 UTC
NVIDIA
Finetune Mistral 7B Using Brev.dev Quick Deploy
In this notebook, we will use NVIDIA's NeMo framework to finetune the Mistral 7B LLM. Finetuning can be done using Brev quick deploy option.
AI
Automatic Speech Recognition
+4
DL
Jupyter Notebook
Language Modeling
NeMo
Resource
23mo
Updated
08/13/2024 UTC
NVIDIA
RIVA ASR Russian LM
Base Russian n-gram LM
Language Modeling
NVIDIA AI
+2
Riva
Russian
Model
>3y
Updated
04/20/2023 UTC
NVIDIA
RIVA ASR French LM
Base French n-gram LM
French
Language Modeling
+2
NVIDIA AI
Riva
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
RIVA ASR Korean LM
Base Korean n-gram LM
Korean
Language Modeling
+2
NVIDIA AI
Riva
Model
>3y
Updated
04/04/2023 UTC
NVIDIA
Riva ASR Italian LM
Base Italian 4-gram LM
Italian
Language Modeling
+1
NVIDIA AI
Model
>3y
Updated
04/11/2023 UTC
NVIDIA
RIVA ASR Hindi LM
Base Hindi n-gram LM
Hindi
Language Modeling
+2
NVIDIA AI
Riva
Model
>3y
Updated
04/04/2023 UTC
24
Select item
24
48
96
192
24
48
96
192
1-24 of 37 items
1
1
2
2
π