Meta
Meta Llama3.2 3B Instruct ONNX INT4 RTX TensorRT Model Optimizer
Model
Meta
Meta Llama3.2 3B Instruct ONNX INT4 RTX TensorRT Model Optimizer

Meta Llama 3.2 3B Instruct INT4 ONNX model is the quantized version of the Meta Llama-3.2-3B-Instruct model, which is an auto-regressive language model that uses an optimized transformer architecture.

NameSizeUpdatedActions
3rd_party_licenses.txt
2.34 KBFebruary 28, 2025 UTC
Acceptable_use_policy.txt
41 BFebruary 28, 2025 UTC
genai_config.json
1.7 KBFebruary 28, 2025 UTC
License.txt
366 BFebruary 28, 2025 UTC
model.onnx
308.75 KBFebruary 28, 2025 UTC
model.onnx_data
2.85 GBFebruary 28, 2025 UTC
notice.txt
121 BFebruary 28, 2025 UTC
Readme.txt
536 BFebruary 28, 2025 UTC
special_tokens_map.json
312 BFebruary 28, 2025 UTC
tokenizer_config.json
55.26 KBFebruary 28, 2025 UTC
tokenizer.json
9.06 MBFebruary 28, 2025 UTC