APOLLO is a single, unified model that can make video and audio together or separately, and it keeps them tightly in sync.
LTX-2 is an open-source model that makes video and sound together from a text prompt, so the picture and audio match in time and meaning.
LiveTalk turns slow, many-step video diffusion into a fast, 4-step, real-time system for talking avatars that listen, think, and respond with synchronized video.
Seedance 1.5 pro is a single model that makes video and sound together at the same time, so lips, music, and actions match naturally.
VABench is a new, all-in-one test that checks how well AI makes videos with matching sound and pictures.