InfiniteTalk: Audio-Driven Full-Body Video Dubbing

InfiniteTalk: Audio-Driven Full-Body Video Dubbing turns a single image and one or more audio tracks into a lip‑synced, full‑body video while preserving the original identity, background, and camera motion. The core of the pipeline is WanInfiniteTalkToVideo, which fuses features from the loaded Wan2.1 I2V model with an audio embedding to drive body and facial motion. Step1 - Load Models prepares the UNet and VAE (UNETLoader, VAELoader) for image‑to‑video synthesis, and loads text and audio backbones (CLIPLoader with CLIPTextEncode for camera/stage prompts, and AudioEncoderLoader with AudioEncoderEncode for voice features). You import the starting frame via LoadImage, then draw masks per character in the MaskEditor to localize motion to each speaker.

FAQ

Sık Sorulan Sorular

Tüm iş akışlarını gör
Character
Cinematic
Image to Video
Lip Sync
Multiple Angles
Portrait
Style Reference
Style Transfer
Text to Video
Video Generation
Video
Showing 30 of 580 templates