Tag
Google highlights PhD intern Khoi's research on synchronizing lip movements with dubbed audio to create natural-looking multi-language videos for YouTube.
Duix.Avatar is a local AI avatar toolkit that generates lip-synced avatar videos from video, voice, and script inputs without sending data to the cloud. It supports eight languages and is available under a community license on GitHub.
NVIDIA Audio2Face-3D generates high-fidelity 3D facial animations from audio, providing accurate lip-sync and emotional expression. It is released as an open-source SDK with pre-trained models and plugins for Maya and Unreal Engine 5.
A discussion about the evolution of AI social apps from text chat to real-time video interfaces, highlighting Mel's multimodal interaction stack and the technical challenges of latency, lip sync, and orchestration.
Vaani is a lip-synced AI dubbing tool for creators, brands, and studios.
Meituan open-sourced the LongCat-Video-Avatar-1.5 model, which supports generating realistic talking videos from a single photo and voice, supports multiple languages and long videos, and outperforms commercial closed-source solutions.
LongCat-Video-Avatar 1.5 is an upgraded open-source framework for audio-driven human video generation with improved lip synchronization, production-ready stability, and efficient 8-step inference.
AI lip-sync tools like sync.so can redraw mouth movements to match dubbed audio, potentially neutralizing the long-standing argument that dubs break immersion due to mismatched mouth movements.