Skip to main content
SearchSearch thousands of GPU-optimized Containers, pretrained Models, SDKs, and Helm charts—ready to accelerate AI, digital twins, and HPC from cloud to edge.
NVIDIA AI Enterprise
NVIDIA AI Enterprise
  • NVIDIA NIM
    NVIDIA NIM
  • NIM Container GPUs
    NIM Container GPUs
  • Use Case
    Use Case
    4
  • NVIDIA Platform
    NVIDIA Platform
    5
  • Industry
    Industry
  • Solution
    Solution
    5
    3
    2
  • Publisher
    Publisher
    6
  • Policy
    Policy
  • Displaying 6 results
    Tacotron2 Speech Synthesis model trained on female English speech
    Model
    Conformer-CTC-Large model for Russian Automatic Speech Recognition, trained on Mozilla Common Voice 10.0 (Russian), Golos (Russian), Russian LibriSpeech (RuLS) and SOVA (RuAudiobooksDevices, RuDevices) datasets.
    Model
    Citrinet 512 model finetuned on Spanish Mozilla CommonVoice and Multilingual LibriSpeech datasets
    Model
    Conformer-Transducer-Large model for Russian Automatic Speech Recognition, trained on Mozilla Common Voice 10.0 (Russian), Golos (Russian), Russian LibriSpeech (RuLS) and SOVA (RuAudiobooksDevices, RuDevices) datasets.
    Model
    This model performs joint intent classification and slot filling, directly from audio input. The model treats the problem as an audio-to-text problem, where the output text is the flattened string representation of the semantics annotation.
    Model
    This collections contains the large version (114M) of the Croatian speech recognition model with a FastConformer encoder and a Hybrid decoder (joint RNNT-CTC loss). The model has a vocab size of 256 and emits text with punctuation and capitalization.
    Model