NVIDIA
Llama-3.1-70b-instruct
Model
NVIDIA
Llama-3.1-70b-instruct

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out).

  • NameSizeUpdatedActions
    chat_template.jinja
    5.68 KBJanuary 15, 2026 UTC
    checksums.blake3
    1.64 KBJanuary 15, 2026 UTC
    config.json
    1.73 KBJanuary 15, 2026 UTC
    generation_config.json
    184 BJanuary 15, 2026 UTC
    hf_quant_config.json
    267 BJanuary 15, 2026 UTC
    model-00001-of-00009.safetensors
    4.65 GBJanuary 15, 2026 UTC
    model-00002-of-00009.safetensors
    4.56 GBJanuary 15, 2026 UTC
    model-00003-of-00009.safetensors
    4.61 GBJanuary 15, 2026 UTC
    model-00004-of-00009.safetensors
    4.61 GBJanuary 15, 2026 UTC
    model-00005-of-00009.safetensors
    4.65 GBJanuary 15, 2026 UTC
    model-00006-of-00009.safetensors
    4.64 GBJanuary 15, 2026 UTC
    model-00007-of-00009.safetensors
    4.61 GBJanuary 15, 2026 UTC
    model-00008-of-00009.safetensors
    4.65 GBJanuary 15, 2026 UTC
    model-00009-of-00009.safetensors
    2.81 GBJanuary 15, 2026 UTC
    model.safetensors.index.json
    202.03 KBJanuary 15, 2026 UTC
    special_tokens_map.json
    325 BJanuary 15, 2026 UTC
    tokenizer_config.json
    66.32 KBJanuary 15, 2026 UTC
    tokenizer.json
    16.41 MBJanuary 15, 2026 UTC
    tool_use_config_v2.json
    41 BJanuary 15, 2026 UTC