Model
Meta Llama 3.2 3B Instruct INT4 ONNX model is the quantized version of the Meta Llama-3.2-3B-Instruct model, which is an auto-regressive language model that uses an optimized transformer architecture.
Use the NGC CLI to download:
Copied!
| Name | Size | Updated | Actions |
|---|---|---|---|
3rd_party_licenses.txt | 2.34 KB | February 28, 2025 UTC | |
Acceptable_use_policy.txt | 41 B | February 28, 2025 UTC | |
genai_config.json | 1.7 KB | February 28, 2025 UTC | |
License.txt | 366 B | February 28, 2025 UTC | |
model.onnx | 308.75 KB | February 28, 2025 UTC | |
model.onnx_data | 2.85 GB | February 28, 2025 UTC | |
notice.txt | 121 B | February 28, 2025 UTC | |
Readme.txt | 536 B | February 28, 2025 UTC | |
special_tokens_map.json | 312 B | February 28, 2025 UTC | |
tokenizer_config.json | 55.26 KB | February 28, 2025 UTC | |
tokenizer.json | 9.06 MB | February 28, 2025 UTC |