Skip to main content
SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
645
383
254
99
96
95
8
4
4
3
2
2
1
1
  • NVIDIA NIM
    NVIDIA NIM
    180
  • NIM Container GPUs
    NIM Container GPUs
    11
    10
    9
    8
    7
    6
    5
    5
    4
    4
    3
    3
    3
    3
    3
    3
    2
    2
    2
    1
    1
  • Use Case
    Use Case
    18
    12
    9
    9
    8
    8
    7
    5
    4
    4
    3
    3
    2
    2
    2
    2
    1
    1
    1
    1
  • NVIDIA Platform
    NVIDIA Platform
    77
    33
    31
    23
    18
    14
    12
    10
    8
    7
    7
    6
    6
    6
    5
    5
    4
    3
    3
    3
    2
    2
    2
    1
    1
    1
    1
    1
    1
    1
  • Industry
    Industry
    37
    15
    14
    13
    12
    8
    8
    7
    6
    6
    6
    4
    4
    4
    3
    3
    2
    1
  • Solution
    Solution
    94
    87
    46
    44
    39
    28
    21
    20
    14
    10
    10
    7
    7
    6
    6
    5
    4
    2
    1
    1
    1
  • Publisher
    Publisher
    616
    3
    2
    2
    2
    2
    2
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
  • Policy
    Policy
    46
  • Displaying 713 results
    DCGM Exporter converts selected DCGM telemetry fields to Prometheus text format. Run one exporter on each GPU node that Prometheus monitors.
    Container
    Validates NVIDIA GPU Operator components
    Container
    Docker containers distributed as part of the TAO Toolkit package
    Container
    Build and Run GPU Accelerated Docker Containers.
    Container
    Deploy and Manage NVIDIA GPU resources in Kubernetes.
    Container
    Provision NVIDIA GPU Driver as a Container
    Container
    cuda
    NVIDIA
    Container registry for CUDA images
    Container
    Triton Inference Server is an open source software that lets teams deploy trained AI models from any framework, from local or cloud storage and on any GPU- or CPU-based infrastructure in the cloud, data center, or embedded devices.
    Container
    Plugin for the Kubernetes Node Feature Discovery for adding GPU node labels.
    Container
    DCGM
    NVIDIA
    Manage and Monitor GPUs in Cluster Environments.
    Container
    PyTorch
    NVIDIA
    PyTorch is a GPU accelerated tensor computational framework. Functionality can be extended with common Python libraries such as NumPy and SciPy. Automatic differentiation is done with a tape-based system at the functional and neural network layer levels.
    Container
    NVIDIA Network Operator Helm Chart provides an easy way to install, configure and manage the lifecycle of NVIDIA Mellanox network operator.
    Helm Chart
    TensorFlow is an open source platform for machine learning. It provides comprehensive tools and libraries in a flexible architecture allowing easy deployment across a variety of platforms and devices.
    Container
    NVIDIA NeMo™ framework Megatron backend supports pre-training, post-training, and reinforcement learning of LLMs and multi-modal generative AI models with state-of-the-art data processing, model training techniques, and flexible deployment options.
    Container
    TensorRT
    NVIDIA
    NVIDIA TensorRT is a C++ library that facilitates high-performance inference on NVIDIA graphics processing units (GPUs). TensorRT takes a trained network and produces a highly optimized runtime engine that performs inference for that network.
    Container
    Llama 3.1 70B-Instruct NIM Production Branch October 2024 (PB 24h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
    Container
    DeepStream SDK delivers a complete streaming analytics toolkit for AI based video and image understanding and multi-sensor processing. This container is for NVIDIA Enterprise GPUs.
    Container
    NVIDIA NIM for GPU accelerated Llama 3.1 8B inference through OpenAI compatible APIs
    Container
    An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
    Container
    This container houses the Llama-3.1-70B-Instruct, which is a multilingual large language model from the Meta Llama 3.1 collection of pretrained and instruction-tuned generative models.
    Container
    CUDA is a parallel computing platform and programming model that enhances computing performance using NVIDIA GPUs. CUDA Deep Learning integrates networking and GPU-accelerated libraries like cuDNN, cuTensor, NCCL, HPC-x, and the CUDA Toolkit.
    Container
    NVIDIA NIM for GPU accelerated Llama-3.1-8B-Instruct inference through OpenAI compatible APIs
    Container
    Parakeet 0.6b CTC en-US NIM delivers accurate English speech-to-text transcription and enables easy-to-use optimized ASR inference for large scale deployments.
    Container
    DOCA Telemetry Service (DTS) is a DOCA Service for collecting and exporting telemetry data. Both predefined and user-defined telemetry counters are available and the export is made by Prometheus, Fluent Bit, Open Telemetry and netflow.
    Container
    ...

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.