NVIDIA
Llama Nemotron Embed 300M v2
Container
NVIDIA
Llama Nemotron Embed 300M v2

World-class multilingual and cross-lingual question-answering retrieval embedding NIM.

Text Embedding NIM

NeMo Retriever Text Embedding NIM (Text Embedding NIM) brings the power of state-of-the-art text embedding models to your applications, offering unparalleled natural language processing and understanding capabilities. You can use Text Retriever NIM for semantic search, Retrieval Augmented Generation (RAG) pipelines, or any application that uses text embeddings. Text Embedding NIM is built on the NVIDIA software platform, incorporating CUDA, TensorRT, and Triton to offer out-of-the-box GPU acceleration.

Getting started with the NIM

Deploying and integrating the NIM is straightforward thanks to our industry standard APIs. Visit the NIM Container page for release documentation, deployment guides and more.

Security Vulnerabilities in Open Source Packages

Please review the Security Scanning tab on NGC to view the latest security scan results.

For certain open-source vulnerabilities listed in the scan results, NVIDIA provides a response in the form of a Vulnerability Exploitability eXchange (VEX) document. The VEX information can be reviewed and downloaded from the Security Scanning tab.

Get Help

NVIDIA Developer Community Forum

For support, Visit the NVIDIA Developer Community Forum

Governing Terms

The NIM container is governed by NVIDIA Agreements | Enterprise Software | NVIDIA Software License Agreement and NVIDIA Agreements | Enterprise Software | Product Specific Terms for AI Product; and the use of this model is governed by the NVIDIA Open Model License.

ADDITIONAL INFORMATION: Llama 3.2 Community License Agreement.

You are responsible for ensuring that your use of NVIDIA AI Foundation Models complies with all applicable laws.

End of Support — "This artifact is no longer supported. NVIDIA strongly recommends artifacts that are up to date and supported"

Publisher
NVIDIA
LicenseNVIDIA proprietary
Latest Tag1.13.0
UpdatedMarch 4, 2026 UTC
Compressed Size2.42 GB
Multinode SupportNo
Multi-Arch SupportYes