NVIDIA
NVIDIA
NeMo Customizer RL Training Container
Container
NVIDIA
NVIDIA
NeMo Customizer RL Training Container

Model customization for NeMo microservices

Sign in to access all content for this ContainerSigning in will also allow download accessSign In

NeMo Customizer RL Training Container

This container provides alignment training functionality for the NeMo Customizer ecosystem. It is specifically designed to run model alignment and customization jobs using Direct Preference Optimization (DPO).

The NeMo Customizer RL image integrates NeMo RL framework with the Customizer training infrastructure, enabling advanced post-training techniques that align language models with human preferences and specific task requirements. It supports distributed training across multiple GPUs and nodes.

Resources

Helm Chart | User Guide

Note: Use, distribution or deployment of this microservice in production requires an NVIDIA AI Enterprise License.

Governing Terms

The software and materials are governed by the NVIDIA Software License Agreement and the Product-Specific Terms for NVIDIA AI Products.

Publisher
NVIDIA
NVIDIA
Latest Tag25.12
UpdatedDecember 16, 2025 UTC
Compressed Size17.15 GB
Multinode SupportNo
Multi-Arch SupportYes

NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.