Tag
Seedance 2.5 launched, enabling 30-second single-pass video generation, audio-reference lip-sync, and API features like async jobs and webhooks. Its standout is billing for failed generations isn't charged and reference videos get lower token rates, though output max is 720p.
Google highlights PhD intern Khoi's research on synchronizing lip movements with dubbed audio to create natural-looking multi-language videos for YouTube.
Duix.Avatar is a local AI avatar toolkit that generates lip-synced avatar videos from video, voice, and script inputs without sending data to the cloud. It supports eight languages and is available under a community license on GitHub.
NVIDIA Audio2Face-3D generates high-fidelity 3D facial animations from audio, providing accurate lip-sync and emotional expression. It is released as an open-source SDK with pre-trained models and plugins for Maya and Unreal Engine 5.
A discussion about the evolution of AI social apps from text chat to real-time video interfaces, highlighting Mel's multimodal interaction stack and the technical challenges of latency, lip sync, and orchestration.
Vaani is a lip-synced AI dubbing tool for creators, brands, and studios.
Meituan open-sourced the LongCat-Video-Avatar-1.5 model, which supports generating realistic talking videos from a single photo and voice, supports multiple languages and long videos, and outperforms commercial closed-source solutions.
LongCat-Video-Avatar 1.5 is an upgraded open-source framework for audio-driven human video generation with improved lip synchronization, production-ready stability, and efficient 8-step inference.
AI lip-sync tools like sync.so can redraw mouth movements to match dubbed audio, potentially neutralizing the long-standing argument that dubs break immersion due to mismatched mouth movements.