NVIDIA
mixtral-8x7b-instruct-v0-1
Container
NVIDIA
mixtral-8x7b-instruct-v0-1

This container houses the Mixtral-8x7B-Instruct-v0.1, a sophisticated, instruction-tuned model from Mistral AI. It uses a Sparse Mixture of Experts (SMoE) architecture, dynamically selecting 2 of its 8 experts per token for greater efficiency.

NIM MetadataTo verify your workspace, run the Verify CLI locally and compare the generated hash with the NGC-published hash shown here. Learn more in the NVIDIA Documentation.
GPU Type
GPU Type
  • GPU Count
    GPU Count
  • Precision
    Precision
  • Profile
    Profile
  • LoRA
    LoRA
  • Engine
    Engine
  • TP
    TP
  • 49 Instances
    Columns
  • Profile
    LoRA
    Engine
    Profile ID
    Actions
    A100 SXM4 80GB20b2:10debf16throughputtensorrt_llm2
    A10G2237:10debf16latencytensorrt_llm8
    B2002901:10defp8latencytensorrt_llm2
    H200 NVL233b:10debf16throughputtensorrt_llm1
    H100 80GB HBM32330:10defp8throughputtensorrt_llm1
    H100 80GB HBM32330:10defp8latencytensorrt_llm2
    B2002901:10debf16throughputtensorrt_llm1
    L40S26b9:10debf16latencytensorrt_llm4
    H200 NVL233b:10defp8latencytensorrt_llm2
    H100 NVL2321:10debf16throughputtensorrt_llm1