SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
1
  • NVIDIA NIM
    NVIDIA NIM
    27
  • NIM Container GPUs
    NIM Container GPUs
    3
    3
    2
    1
    1
    1
  • Use Case
    Use Case
    3
    3
  • NVIDIA Platform
    NVIDIA Platform
    1
  • Industry
    Industry
  • Solution
    Solution
    3
    2
  • Publisher
    Publisher
    14
    3
    3
    1
    1
    1
    1
    1
    1
    1
  • Policy
    Policy
  • Displaying 27 results
    The MiniMax-M2.5 NIM Container is a deployable inference container for serving MiniMax-M2.5, a third-party text generation model optimized for complex agentic tasks including software engineering, tool use, search.
    Container
    This container houses GLiNER PII, which detects and classifies a broad range of Personally Identifiable Information (PII) and Protected Health Information (PHI) in structured and unstructured text.
    Container
    Nemotron-3-Super-120B-A12B is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.
      Container
      Gemma 4 26B A4B IT is a Google multimodal instruction-tuned model packaged as an NVIDIA NIM container for deployment through NVIDIA NGC as a Downloadable NIM.
        Container
        Qwen3.5-35B-A3B is a multimodal vision-language Mixture-of-Experts model designed for native multimodal agent applications, supporting text, image, and video inputs.
          Container
          Gemma 4 31B IT model which, is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output.
          Container
          Step 3.5 Flash is a sparse Mixture-of-Experts (MoE) large language model developed by StepFun, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency
          Container
          The Qwen3.6-35B-A3B NIM Container is a deployable inference container for serving Qwen3.6-35B-A3B, a third-party multimodal Mixture of Experts model capable of processing text, image, and video inputs for text generation.
            Container
            Moonshot AI
            kimi-k2.5-Turbo
            This turbo container houses the Kimi K2.5 model which is an open-source, native multimodal agentic model built through continual pretraining on approximately 15 trillion mixed visual and text tokens atop Kimi-K2-Base.
              Container
              zai-org
              glm-5.2
              This container houses GLM-5.2, a flagship long-context large language model for agentic engineering and advanced reasoning, packaged as an NVIDIA NIM for staging.
                Container
                Nemotron Nano V3 Omni is a multi-modal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
                  Container
                  Z.Ai
                  GLM-5
                  GLM-5 is a next-generation large language model targeting complex systems engineering and long-horizon agentic tasks.
                    Container
                    Nemotron Content Safety Reasoning 4B is a Large Language Model (LLM) classifier designed to function as a dynamic and adaptable guardrail for content safety and dialogue moderation (topic-following).
                    Container
                    The NVIDIA Ising Calibration 1 NIM houses the NVIDIA-Ising-Calibration-1-35B-A3B-BF16 model, which is a purpose-built Mixture-of-Experts vision-language model (MoE VLM) built on Qwen3.5-35B-A3B,
                      Container
                      NVIDIA
                      NVIDIA
                      GLM-5.1
                      This container houses GLM-5.1, which is a next-generation flagship model for agentic engineering with significantly stronger coding capabilities than its predecessor GLM-5. The model achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5
                        Container
                        This container houses the model MiMo-V2-Flash.
                        Container
                        The Nemotron 3.5 Content Safety NIM container packages NVIDIA's small language model (SLM) that uses Google's Gemma-3-4B-it as the base and is fine-tuned by NVIDIA on multimodal, multilingual, and reasoning-oriented content-safety datasets.
                          Container
                          The DiffusionGemma-4-26B-A4B-IT model is an open-weights multimodal generative model developed by Google DeepMind that processes text, image, and video inputs to produce text output via discrete diffusion.
                            Container
                            The Qwen3.6-27B NIM Container is a deployable inference container for serving Qwen3.6-27B, a third-party multimodal dense model capable of processing text, image, and video inputs for text generation.
                              Container
                              This NIM container houses the Nemotron 3 Content Safety model which, is a small language model (SLM) that uses Google's Gemma-3-4B-it as the base and is fine-tuned by NVIDIA on multimodal and multilingual content-safety related datasets.
                                Container
                                Deepseek AI
                                DeepSeek-V4-Pro
                                The DeepSeek-V4-Pro Container is a deployable inference container for serving DeepSeek-V4-Pro, a third-party sparse Mixture-of-Experts language model for reasoning, coding, and agentic tasks.
                                  Container
                                  Nemotron-3-Super-120B-A12B is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.
                                    Container
                                    Stepfun-AI
                                    Step 3.7 Flash
                                    The Step-3.7-Flash NIM is a Downloadable NIM container for deploying Step-3.7-Flash, a StepFun vision-language model built on Step 3.5 Flash with additional vision capability for native multimodal, agentic, and coding-related use cases.
                                    Container
                                    Thinking Machines Lab
                                    Inkling
                                    This NIM container houses Inkling model, which is a 66-layer decoder-only transformer with a sparse Mixture-of-Experts (MoE) feed-forward backbone, featuring 975B total parameters and 41B active parameters
                                      Container