SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
51
35
11
10
3
2
1
1
1
1
1
  • NVIDIA NIM
    NVIDIA NIM
    13
  • NIM Container GPUs
    NIM Container GPUs
  • Use Case
    Use Case
    59
    37
    23
    23
    22
    21
    17
    16
    10
    7
    7
    6
    6
    6
    5
    5
    5
    4
    4
    4
    4
    2
    2
    2
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
  • NVIDIA Platform
    NVIDIA Platform
    48
    30
    26
    19
    17
    15
    15
    13
    10
    10
    9
    7
    7
    6
    6
    5
    4
    4
    4
    3
    3
    2
    2
    1
    1
    1
    1
    1
    1
    1
    1
    1
  • Industry
    Industry
    58
    25
    25
    23
    18
    16
    15
    12
    9
    9
    8
    6
    6
    5
    5
    5
    5
    5
    2
    1
  • Solution
    Solution
    397
    270
    258
    238
    176
    169
    112
    111
    103
    55
    49
    46
    40
    16
    16
    14
    12
    11
    11
    9
    8
    4
    3
    3
    2
    1
    1
  • Publisher
    Publisher
    297
    12
    10
    8
    6
    5
    3
    2
    1
    1
    1
    1
    1
    1
    1
    1
  • Policy
    Policy
    9
  • Displaying 397 results
    Docker containers distributed as part of the TAO Toolkit package
    Container
    NVIDIA
    NVIDIA
    PyTorch
    PyTorch is a GPU accelerated tensor computational framework. Functionality can be extended with common Python libraries such as NumPy and SciPy. Automatic differentiation is done with a tape-based system at the functional and neural network layer levels.
    Container
    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs.
    Container
    DeepStream SDK delivers a complete streaming analytics toolkit for AI based video and image understanding and multi-sensor processing. This container is for NVIDIA Enterprise GPUs.
    Container
    NVIDIA Developer Program
    GenMol is a masked diffusion model trained on molecular SAFE representations for fragment-based molecule generation, which can serve as a generalist model for various drug discovery tasks.
    Container
    NVIDIA
    NVIDIA
    vLLM
    vLLM is a fast and easy-to-use library for LLM inference and serving. The NVIDIA vLLM NGC Container is optimized for GPU acceleration, and contains a validated set of libraries that enable and optimize GPU performance.
    Container
    An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
    Container
    CUDA is a parallel computing platform and programming model that enhances computing performance using NVIDIA GPUs. CUDA Deep Learning integrates networking and GPU-accelerated libraries like cuDNN, cuTensor, NCCL, HPC-x, and the CUDA Toolkit.
    Container
    NVIDIA Developer Program
    Parakeet 0.6b CTC en-US NIM delivers accurate English speech-to-text transcription and enables easy-to-use optimized ASR inference for large scale deployments.
    Container
    NVIDIA Developer Program
    NVIDIA
    NVIDIA
    NV-CLIP
    NV-CLIP NIM microservice for multimodal embeddings model for image and text
    Container
    NVIDIA Developer Program
    Diffdock predicts the 3D structure of the interaction between a molecule and a protein.
    Container
    This is the RAG server container used as part of the NVIDIA AI Blueprint for RAG and used to orchestrate the end to end RAG pipeline.
    Container
    NVIDIA Developer Program
    NVIDIA
    NVIDIA
    MolMIM
    MolMIM is a transformer-based model developed by NVIDIA for controlled small molecule generation.
    Container
    The Dynamo vLLM runtime image is a containerized build of Dynamo + vLLM which serves as the base runtime environment for vLLM based inference with Dynamo's distributed inference framework.
    Container
    This is the Ingestor server container used as part of the NVIDIA RAG Blueprint and used to orchestrate the end to end Ingestion.
    Container
    The Dynamo TensorRT-LLM runtime image is a containerized build of Dynamo + TensorRT-LLM which serves as the base runtime environment for tensorrt-llm based inference with Dynamo's distributed inference framework.
    Container
    NVIDIA
    NVIDIA
    JAX
    JAX is a framework for high-performance numerical computing and machine learning research. It includes Numpy-like APIs, automatic differentiation, XLA acceleration and simple primitives for scaling across GPUs and supports an ecosystem of libraries.
    Container
    A comprehensive Helm chart for deploying the NVIDIA Dynamo operator and its dependencies
    Helm Chart
    NVIDIA
    NVIDIA
    PyG
    PyG (PyTorch Geometric) is a library built upon PyTorch to easily write and train Graph Neural Networks (GNNs) for a wide range of applications related to structured data.
    Container
    Helm chart for NIM Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
    Helm Chart
    Build a Video Search and Summarization Agent Ingest massive volumes of live or archived videos and extract insights for summarization and interactive Q&A
    Container
    The Dynamo SGLang runtime image is a containerized build of Dynamo + SGLang which serves as the base runtime environment for sglang based inference with Dynamo's distributed inference framework.
    Container
    TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs.
    Container
    kubernetes-operator is a container that runs as part of the Dynamo cloud platform. Dynamo cloud is a kubernetes platform for deploying and managing inference services. This container manages the lifecycle of Dynamo inference deployments in kubernetes.
    Container
    ...