Skip to main content
SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
  • NVIDIA NIM
    NVIDIA NIM
  • NIM Container GPUs
    NIM Container GPUs
  • Use Case
    Use Case
    20
    13
  • NVIDIA Platform
    NVIDIA Platform
    35
    19
    2
  • Industry
    Industry
  • Solution
    Solution
    35
    16
    6
    4
  • Publisher
    Publisher
    82
  • Policy
    Policy
  • Displaying 82 results
    English Conformer ASR model trained on ASR set 6.0
    Model
    Arabic (ar-AR) Conformer ASR model trained on ASR set 2.0
    Model
    English Conformer ASR model for en-US
    Model
    Citrinet 1024model trained on ASR Set dataset
    Model
    QuartzNet is a Jasper-like network that uses separable convolutions and larger filter sizes. It has comparable accuracy to Jasper while having much fewer parameters. This particular model has 15 blocks each repeated 5 times.
    Model
    Speech To Text (STT) model based on QuartzNet for recognizing Spanish speech.
    Model
    Citrinet-1024 model with kernel scaling factor (gamma) of 25%, which has been trained on the open-source Aishell-2 Mandarin Chinese corpus.
    Model
    Speech To Text (STT) model based on QuartzNet for recognizing Russian speech.
    Model
    Speech to Text Citrinet models for English.
    Model
    English en-US Citrinet ASR model trained on RIVA ASR set
    Model
    Citrinet 512 model trained on Aishell-2 Mandarin corpus
    Model
    QuartzNet is a Jasper-like network that uses separable convolutions and larger filter sizes. It has comparable accuracy to Jasper while having much fewer parameters. This particular model has 15 blocks each repeated 5 times.
    Model
    Conformer-Transducer-Large model for Mandarin Automatic Speech Recognition, Trained on Aishell-2 Mandarin Chinese corpus.
    Model
    Conformer-CTC-XLarge model for English Automatic Speech Recognition, Trained on NeMo ASRSET
    Model
    Jasper models are end-to-end neural automatic speech recognition (ASR) models that transcribe segments of audio to text.
    Model
    German Conformer ASR model trained on RIVA ASR set
    Model
    Mandarin (zh-CN) Conformer ASR model trained on ASR set 5.0
    Model
    Speech To Text (STT) model based on QuartzNet for recognizing French speech.
    Model
    Hindi Conformer ASR model trained on RIVA ASR set
    Model
    Russian Conformer ASR model trained on RIVA ASR set
    Model
    Japanese Conformer ASR model trained on RIVA ASR set
    Model
    French Conformer ASR model trained on RIVA ASR set
    Model
    Spanish Conformer ASR model trained on ASR set 3.0
    Model
    Conformer-CTC-Medium model for Hindi Automatic Speech Recognition, Trained on ULCA Hindi Labelled Dataset.
    Model

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.