NVIDIA AI for Media LipSync syncs lip movements in a video to match the provided speech audio, enhancing realism and accuracy in speech animation. Takes an audio and video input, generating a synchronized output video.
What Is NVIDIA NIM?
NVIDIA NIM™, part of NVIDIA AI Enterprise, is a set of easy-to-use microservices designed for secure, reliable deployment of high performance AI model inferencing across clouds, data centers and workstations. Supporting a wide range of AI models, including open-source and NVIDIA AI Foundation and custom models, it ensures seamless, scalable AI inferencing, on-premises or in the cloud, leveraging industry standard APIs.
NVIDIA Lipsync NIM is an AI-powered service that synchronizes lip movements in videos with an input spoken audio, creating naturally synchronized speech animations.
NVIDIA NIM offers prebuilt containers for computer vision models. Each NIM consists of a container and a model and uses a CUDA-accelerated runtime for all NVIDIA GPUs, with special optimizations available for many configurations. Whether on-premises or in the cloud, NIM is the fastest way to achieve accelerated inference at scale.
Getting Started with NVIDIA NIM
Deploying and integrating NVIDIA NIM is straightforward thanks to our industry standard APIs. Visit the NVIDIA Lipsync NIM page for release documentation, deployment guides and more.
Governing Terms
Use of the NIM container is governed by the NVIDIA Software License Agreement and Product-Specific Terms for NVIDIA AI Products. Use of this model is governed by the NVIDIA Open Model License.
You are responsible for ensuring that your use of NVIDIA AI Foundation Models complies with all applicable laws.
Get Help
NVIDIA Developer Community Forum
For support, Visit the NVIDIA Developer Community Forum
End of Support — "This artifact is no longer supported. NVIDIA strongly recommends artifacts that are up to date and supported"