Model
Nemotron-Mini-4B Instruct model is for generating responses for roleplaying, retrieval augmented generation, and function calling. It is a small language model optimized through distillation, pruning and quantization for speed and on-device deployment.
Use the NGC CLI to download:
Copied!
| Name | Size | Updated | Actions |
|---|---|---|---|
3rd_party_licenses.txt | 2.34 KB | February 28, 2025 UTC | |
genai_config.json | 1.62 KB | February 28, 2025 UTC | |
License.txt | 351 B | February 28, 2025 UTC | |
model.onnx | 385.09 KB | February 28, 2025 UTC | |
model.onnx_data | 4.19 GB | February 28, 2025 UTC | |
quantization_log.txt | 3.54 KB | February 28, 2025 UTC | |
quantization_pip.txt | 11.54 KB | February 28, 2025 UTC | |
readme.txt | 0 B | February 28, 2025 UTC | |
Readme.txt | 536 B | February 28, 2025 UTC | |
special_tokens_map.json | 292 B | February 28, 2025 UTC | |
tokenizer_config.json | 184.11 KB | February 28, 2025 UTC | |
tokenizer.json | 33.2 MB | February 28, 2025 UTC |