NVIDIA
Deepseek-R1-Distill-Llama-8B
Container
NVIDIA
Deepseek-R1-Distill-Llama-8B

NVIDIA NIM for GPU accelerated DeepSeek-R1-Distill-Llama-8B inference through OpenAI compatible APIs

NIM MetadataTo verify your workspace, run the Verify CLI locally and compare the generated hash with the NGC-published hash shown here. Learn more in the NVIDIA Documentation.
GPU Type
GPU Type
  • GPU Count
    GPU Count
  • Precision
    Precision
  • Profile
    Profile
  • LoRA
    LoRA
  • Engine
    Engine
  • TP
    TP
  • 26 Instances
    Columns
  • Profile
    LoRA
    Engine
    Profile ID
    Actions
    H100 NVL2321:10defp8throughputtensorrt_llm1
    H100 80GB HBM32330:10defp8throughputtensorrt_llm1
    H2002335:10defp8latencytensorrt_llm2
    H100 80GB HBM32330:10debf16latencytensorrt_llm2
    A10G2237:10debf16latencytensorrt_llm4
    L40S26b9:10debf16latencytensorrt_llm2
    H100 NVL2321:10defp8latencytensorrt_llm2
    bf16generictensorrt_llm2
    H100 80GB HBM32330:10debf16throughputtensorrt_llm1
    bf16genericvllm1