Google
CodeGemma-7B-IT-INT4-RTX
Resource
Google
CodeGemma-7B-IT-INT4-RTX

CodeGemma is a collection of lightweight open code models built on top of Gemma. CodeGemma models are text-to-text decoder-only models. This is a 7 billion parameter instruction-tuned varient for code chat and instruction.

  • Model Overview

    Description:

    CodeGemma is a collection of lightweight open code models built on top of Gemma. CodeGemma models are text-to-text decoder-only models. This is a 7 billion parameter instruction-tuned varient for code chat and instruction.

    Terms of use:

    Use of this model is governed by the NVIDIA AI Foundation Models Community License. ADDITIONAL INFORMATION: Gemma Terms of Use and Google Prohibited Use Policy.

    References(s):

    Model Architecture:

    Architecture Type: Transformer

    Input:

    Input Format: Text

    Input Parameters: None

    Output:

    Output Format: Text

    Output Parameters: None

    Software Integration:

    Supported Hardware Platform(s): RTX 4090

    Supported Operating System(s): Windows

    Inference:

    TRT-LLM Inference Engine
    Windows Setup with TRT-LLM

    Test Hardware:

    RTX 4090

    Publisher
    Google
    Latest Version1.0
    UpdatedApril 9, 2024 UTC
    Compressed Size6.22 GB
    Labels

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.