Google
Google
Gemma-2b-Instruct-ONNX-INT4-RTX
Model
Google
Google
Gemma-2b-Instruct-ONNX-INT4-RTX

The NVIDIA Gemma-2b-it INT4 ONNX model is the quantized version of the Google Gemma-2b-it model which is a text-to-text, decoder-only large language models, available in English, with open weights, pre-trained variants, and instruction-tuned variants.

NameSizeUpdatedActions
3rd_party_licenses.txt
2.34 KBFebruary 28, 2025 UTC
genai_config.json
1.61 KBFebruary 28, 2025 UTC
License.txt
287 BFebruary 28, 2025 UTC
model.onnx
197.71 KBFebruary 28, 2025 UTC
model.onnx_data
2.91 GBFebruary 28, 2025 UTC
Notice.txt
96 BFebruary 28, 2025 UTC
Prohibited_Use_Policy.txt
41 BFebruary 28, 2025 UTC
readme.txt
0 BFebruary 28, 2025 UTC
Readme.txt
536 BFebruary 28, 2025 UTC
special_tokens_map.json
670 BFebruary 28, 2025 UTC
tokenizer_config.json
41.37 KBFebruary 28, 2025 UTC
tokenizer.json
16.71 MBFebruary 28, 2025 UTC
tokenizer.model
4.04 MBFebruary 28, 2025 UTC

NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.