Model
NVIDIA
Llama-3.3-Nemotron-Super-49B-v1.5Llama-3.3-Nemotron-Super-49B-v1.5 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct
By using the Model(s), you agree to the Model License(s).
Use the NGC CLI to download:
Copied!
| Name | Size | Updated | Actions |
|---|---|---|---|
accuracy_chart.png | 178.59 KB | August 4, 2026 UTC | |
BIAS.md | 208 B | August 4, 2026 UTC | |
block_config.py | 4.25 KB | August 4, 2026 UTC | |
checksums.blake3 | 4.49 KB | August 4, 2026 UTC | |
config.json | 36.21 KB | August 4, 2026 UTC | |
configuration_decilm.py | 2.51 KB | August 4, 2026 UTC | |
EXPLAINABILITY.md | 2.93 KB | August 4, 2026 UTC | |
generation_config.json | 176 B | August 4, 2026 UTC | |
llama_nemotron_toolcall_parser_no_streaming.py | 23.94 KB | August 4, 2026 UTC | |
model-00001-of-00021.safetensors | 4.64 GB | August 4, 2026 UTC | |
model-00002-of-00021.safetensors | 4.63 GB | August 4, 2026 UTC | |
model-00003-of-00021.safetensors | 4.66 GB | August 4, 2026 UTC | |
model-00004-of-00021.safetensors | 4.57 GB | August 4, 2026 UTC | |
model-00005-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00006-of-00021.safetensors | 4.66 GB | August 4, 2026 UTC | |
model-00007-of-00021.safetensors | 4.63 GB | August 4, 2026 UTC | |
model-00008-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00009-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00010-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00011-of-00021.safetensors | 4.66 GB | August 4, 2026 UTC | |
model-00012-of-00021.safetensors | 4.63 GB | August 4, 2026 UTC | |
model-00013-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00014-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00015-of-00021.safetensors | 4.44 GB | August 4, 2026 UTC | |
model-00016-of-00021.safetensors | 4.44 GB | August 4, 2026 UTC | |
model-00017-of-00021.safetensors | 4.59 GB | August 4, 2026 UTC | |
model-00018-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00019-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00020-of-00021.safetensors | 4.34 GB | August 4, 2026 UTC | |
model-00021-of-00021.safetensors | 3.27 GB | August 4, 2026 UTC | |
model.safetensors.index.json | 45.69 KB | August 4, 2026 UTC | |
modeling_decilm.py | 76.74 KB | August 4, 2026 UTC | |
nemotron_deci.py | 13.59 KB | August 4, 2026 UTC | |
PRIVACY.md | 514 B | August 4, 2026 UTC | |
README.md | 17.32 KB | August 4, 2026 UTC | |
SAFETY_SECURITY.md | 920 B | August 4, 2026 UTC | |
special_tokens_map.json | 439 B | August 4, 2026 UTC | |
tokenizer_config.json | 64.27 KB | August 4, 2026 UTC | |
tokenizer.json | 16.41 MB | August 4, 2026 UTC | |
tool_use_config_v2.json | 89 B | August 4, 2026 UTC | |
transformers_4_44_2__activations.py | 8 KB | August 4, 2026 UTC | |
transformers_4_44_2__cache_utils.py | 63.08 KB | August 4, 2026 UTC | |
transformers_4_44_2__configuration_llama.py | 10.81 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_attn_mask_utils.py | 20.83 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_flash_attention_utils_backward_compat.py | 15.38 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_outputs.py | 109.94 KB | August 4, 2026 UTC | |
transformers_4_44_2__modeling_rope_utils.py | 27.41 KB | August 4, 2026 UTC | |
transformers_4_44_2__pytorch_utils.py | 666 B | August 4, 2026 UTC | |
variable_cache.py | 6.06 KB | August 4, 2026 UTC |