Model
NVIDIA
Llama-3.3-Nemotron-Super-49B-v1.5Llama-3.3-Nemotron-Super-49B-v1.5 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct
By using the Model(s), you agree to the Model License(s).
Use the NGC CLI to download:
Copied!
| Name | Size | Updated | Actions |
|---|---|---|---|
block_config.py | 4.25 KB | August 4, 2026 UTC | |
chat_template.jinja | 3.7 KB | August 4, 2026 UTC | |
checksums.blake3 | 2.83 KB | August 4, 2026 UTC | |
config.json | 47.05 KB | August 4, 2026 UTC | |
configuration_decilm.py | 2.51 KB | August 4, 2026 UTC | |
generation_config.json | 176 B | August 4, 2026 UTC | |
hf_quant_config.json | 267 B | August 4, 2026 UTC | |
llama_nemotron_toolcall_parser_no_streaming.py | 24.28 KB | August 4, 2026 UTC | |
model-00001-of-00007.safetensors | 4.65 GB | August 4, 2026 UTC | |
model-00002-of-00007.safetensors | 4.65 GB | August 4, 2026 UTC | |
model-00003-of-00007.safetensors | 4.61 GB | August 4, 2026 UTC | |
model-00004-of-00007.safetensors | 4.65 GB | August 4, 2026 UTC | |
model-00005-of-00007.safetensors | 4.6 GB | August 4, 2026 UTC | |
model-00006-of-00007.safetensors | 3.83 GB | August 4, 2026 UTC | |
model-00007-of-00007.safetensors | 1.96 GB | August 4, 2026 UTC | |
model.safetensors.index.json | 157.08 KB | August 4, 2026 UTC | |
modeling_decilm.py | 76.74 KB | August 4, 2026 UTC | |
nemotron_deci.py | 13.59 KB | August 4, 2026 UTC | |
special_tokens_map.json | 439 B | August 4, 2026 UTC | |
tokenizer_config.json | 65.19 KB | August 4, 2026 UTC | |
tokenizer.json | 16.41 MB | August 4, 2026 UTC | |
tool_use_config_v2.json | 89 B | August 4, 2026 UTC | |
transformers_4_44_2__activations.py | 8 KB | August 4, 2026 UTC | |
transformers_4_44_2__cache_utils.py | 63.08 KB | August 4, 2026 UTC | |
transformers_4_44_2__configuration_llama.py | 10.81 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_attn_mask_utils.py | 20.83 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_flash_attention_utils_backward_compat.py | 15.38 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_outputs.py | 109.94 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_rope_utils.py | 27.41 KB | August 4, 2026 UTC | |
transformers_4_44_2__pytorch_utils.py | 666 B | August 4, 2026 UTC | |
variable_cache.py | 6.06 KB | August 4, 2026 UTC |