Skip to main content
SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
74
59
53
4
1
1
1
  • NVIDIA NIM
    NVIDIA NIM
    108
  • NIM Container GPUs
    NIM Container GPUs
    7
    5
    3
    3
    2
    1
    1
    1
    1
    1
  • Use Case
    Use Case
    8
    5
    5
    5
    4
    3
    2
    1
    1
    1
    1
    1
    1
    1
    1
  • NVIDIA Platform
    NVIDIA Platform
    4
    3
    2
    1
    1
  • Industry
    Industry
    13
    2
    1
    1
    1
    1
    1
  • Solution
    Solution
    11
    6
    6
    3
    1
    1
    1
  • Publisher
    Publisher
    100
    3
    2
    2
    1
    1
    1
    1
    1
    1
    1
    1
    1
    1
  • Policy
    Policy
    19
  • Displaying 118 results
    Llama 3.1 70B-Instruct NIM Production Branch October 2024 (PB 24h2) offers a 9-month lifecycle for API stability, with monthly patches for high and critical software vulnerabilities.
    Container
    NVIDIA NIM for GPU accelerated Llama 2 70B inference through OpenAI compatible APIs
    Container
    NVIDIA NIM for GPU accelerated Mistral-NeMo-12B-Instruct inference through OpenAI compatible APIs
    Container
    Hive’s Deepfake Image Detection model analyzes images and returns a confidence score on how likely the image contains a deepfake.
    Container
    NVIDIA NIM for GPU accelerated Llama 3.1 8B inference through OpenAI compatible APIs
    Container
    An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
    Container
    This container houses the Llama-3.1-70B-Instruct, which is a multilingual large language model from the Meta Llama 3.1 collection of pretrained and instruction-tuned generative models.
    Container
    NVIDIA NIM for GPU accelerated CodeLlama-70B inference through OpenAI compatible APIs
    Container
    NVIDIA NIM for GPU accelerated Llama-3.1-8B-Instruct inference through OpenAI compatible APIs
    Container
    NV-CLIP
    NVIDIA
    NV-CLIP NIM microservice for multimodal embeddings model for image and text
    Container
    NVIDIA NIM for GPU accelerated Llama-3.1-Nemotron-70B-Instruct inference through OpenAI compatible APIs
    Container
    RIVA Parakeet 1.1b RNNT Multilingual ASR NIM delivers accurate speech-to-text transcription for 25 languages
    Container
    This container houses the Llama-3.3-Nemotron-Super-49B-v1.5, which is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct
    Container
    Please add descriptionNVIDIA NIM for GPU accelerated Mixtral-8x7B-Instruct-v0.1 inference through OpenAI compatible APIs
    Container
    NVIDIA NIM for GPU accelerated Llama 3.1 405B inference through OpenAI compatible APIs
    Container
    Natural and expressive voices in multiple languages. For voice agents and brand ambassadors.
    Container
    Reranking NIM optimized for providing a logit score that represents how relevant a document(s) is to a given query, fine-tuned for multilingual and cross-lingual text question-answering retrieval.
    Container
    Helm chart for NIM Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment
    Helm Chart
    Parakeet 1.1b CTC en-US ASR NIM delivers accurate English speech-to-text transcription and enables easy-to-use optimized ASR inference for large scale deployments.
    Container
    FLUX.1-schnell is a distilled image generation model, producing high quality images at fast speeds.
    Container
    NVIDIA NIM for GPU accelerated Llama 2 13B inference through OpenAI compatible APIs
    Container
    Hives AI Generated Image Detection model analyzes images and returns a confidence score on how likely the image is AI generated or modified in some way.
    Container
    Nemotron-3-Ultra-550B-A55B NIM container packages NVIDIA's large language model featuring a hybrid Latent Mixture-of-Experts (LatentMoE) architecture with Multi-Token Prediction (MTP) layers.
    Container
    NVIDIA NIM for GPU accelerated Llama 3.1 8B inference through OpenAI compatible APIs
    Container

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.