SkyReels-V4 is a single, unified model that makes videos and matching sounds together, while also letting you fix or change parts of a video.
DreamID-Omni is one model that can create, edit, and animate human-centered videos with matching voices, all in sync.
Ex-Omni is a new open-source AI system that can understand text or speech and then talk back while moving a 3D face in sync with the voice.
APOLLO is a single, unified model that can make video and audio together or separately, and it keeps them tightly in sync.
LTX-2 is an open-source model that makes video and sound together from a text prompt, so the picture and audio match in time and meaning.
LiveTalk turns slow, many-step video diffusion into a fast, 4-step, real-time system for talking avatars that listen, think, and respond with synchronized video.
Seedance 1.5 pro is a single model that makes video and sound together at the same time, so lips, music, and actions match naturally.
VABench is a new, all-in-one test that checks how well AI makes videos with matching sound and pictures.