VEED's Lip Sync 2.0 lands on fal
By Steven Van ·
The model resyncs audio and lip movement on up to 4K, 10-minute clips via fal's inference API, no training required.
fal now hosts VEED's Lip Sync 2.0 model, which swaps the audio on any video and matches lip movement, emotion, style and timing without training or fine-tuning. It handles footage up to 4K and clips up to 10 minutes, with a controllable sync strength setting.
The model is available now through fal's inference API, as announced on X.
fal
A generative media platform and inference API for running fast image, video, and audio AI models at scale.
View fal →