AI Model

Wan ComfyUI Workflows

Run the open-source Wan video model in ComfyUI: text-to-video, image-to-video, and motion control, on Comfy Cloud or your own GPU.

61 workflows

The workflows

61 Wan ComfyUI Workflows

1 image input Split Stack - Qwen Multiangle + Wan 2.2

Wan2.2
RobRob
Video
Wan 2.2 14B Text to Video

Wan 2.2 14B Text to Video

Wan2.2
ComfyUIComfyUI
Text to Video
Video

Image to Video (New)

Wan
ComfyUIComfyUI
Image to Video
Video

Wan ATI Trajectory Control

Wan ATI
Jo Z.Jo Z.
Video
Image to Video

Wan2.2 Animate: Automatic Character Replacement

wan2.2 Animate
PurzPurz
Video
Replacement
Video Edit

Audio-Driven Character Lip Sync Video

Wan2.1 InfiniteTalk
RobRob
Audio to Video
Lip Sync
Video

Wan 2.2 14B Image to Video

Wan
ComfyUIComfyUI
Image to Video
Video
1.2 Starter - Image to Video

1.2 Starter - Image to Video

Wan2.2
ComfyUIComfyUI
Image
Wan2.6: Image to Video

Wan2.6: Image to Video

Wan2.6
ComfyUIComfyUI
Image to Video
Video
Partner Nodes
Wan2.1 SCAIL

Wan2.1 SCAIL

Wan2.1 SCAIL
ComfyUIComfyUI
Pose transfer
Video
Wan2.2 Animate, Character Animation and Replacement

Wan2.2 Animate, Character Animation and Replacement

Wan
ComfyUIComfyUI
Video
Image to Video

Multi-Keyframe Video Stitching

Wan2.2
ComfyUIComfyUI
FLF2V
Image to Video

Inflation Lora Character Expansion Effect Video

wan2.2 Animate
Video
Wan2.5: Image to Video

Wan2.5: Image to Video

Wan2.5
ComfyUIComfyUI
Image to Video
Video
Partner Nodes
Wan 2.2 14B First-Last Frame to Video - After
Wan 2.2 14B First-Last Frame to Video - Before

Wan 2.2 14B First-Last Frame to Video

Wan2.2
ComfyUIComfyUI
FLF2V
Video
Wan2.1 VACE Control Video - After
Wan2.1 VACE Control Video - Before

Wan2.1 VACE Control Video

Wan2.1
ComfyUIComfyUI
Video to Video
Video
Wan 2.2 5B Video Generation

Wan 2.2 5B Video Generation

Wan2.2
ComfyUIComfyUI
Text to Video
Video

Wan ATI Motion Control

Wan ATI
RobRob
Video
Wan2.5: Text to Image

Wan2.5: Text to Image

Wan2.5
ComfyUIComfyUI
Text to Image
Image
Partner Nodes
Wan 2.2 14B Fun Control

Wan 2.2 14B Fun Control

Wan2.2
ComfyUIComfyUI
Video to Video
Video
ControlNet
Wan2.2-S2V Audio-Driven Video Generation

Wan2.2-S2V Audio-Driven Video Generation

Wan2.2
ComfyUIComfyUI
Video
InfiniteTalk: Audio-Driven Full-Body Video Dubbing

InfiniteTalk: Audio-Driven Full-Body Video Dubbing

Wan2.1 InfiniteTalk
ComfyUIComfyUI
Audio to Video
Lip Sync
Video
Wan 2.1 ControlNet - After
Wan 2.1 ControlNet - Before

Wan 2.1 ControlNet

Wan2.1
ComfyUIComfyUI
Video to Video
Video

AIrt MAchIne

Wan2.2
FLF2V
Text to Image
Wan 2.2 14B Fun Inp

Wan 2.2 14B Fun Inp

Wan2.2
ComfyUIComfyUI
FLF2V
Video
Wan 2.1 Image to Video

Wan 2.1 Image to Video

Wan2.1
ComfyUIComfyUI
Text to Video
Video
Wan2.1 VACE Inpainting - After
Wan2.1 VACE Inpainting - Before

Wan2.1 VACE Inpainting

Wan2.1
ComfyUIComfyUI
Inpainting
Video

Video to Seamless Loop Converter

Wan2.1 VACE
SirolimSirolim
Video Edit
Wan 2.1 Text to Video

Wan 2.1 Text to Video

Wan2.1
ComfyUIComfyUI
Text to Video
Video
Wan2.6: Text to Video

Wan2.6: Text to Video

Wan2.6
ComfyUIComfyUI
Text to Video
Video
Partner Nodes
Showing 30 of 61 templates

About This AI Model

What is the Wan model?

If you are comparing Wan ComfyUI workflows, the first thing worth knowing is that Wan is genuinely open. It is Alibaba Tongyi Lab's open-source video family, released under Apache 2.0, which is why you can run it locally on your own hardware instead of renting it by the second. The workflows on this page are the ready-made graphs for doing exactly that: load a Wan model, give it a prompt or an image, and generate a short clip.

Wan covers the jobs people actually reach for. Text-to-video turns a written prompt into motion, image-to-video animates a still you provide, and the newer S2V variants drive a portrait from an audio clip with lip-sync. Across the family you get high temporal consistency, bilingual prompts, and multiple aspect ratios at 24 frames per second, with clips that typically land around 81 frames, roughly three and a half seconds.

Versions matter when you pick a workflow. Wan 2.1 ships in a light 1.3B size that runs on about 8GB of VRAM and a 14B size that wants 24GB or more, both at 480p and 720p, plus Fun variants for camera control and depth, pose, or canny conditioning. Wan 2.2 moves to a mixture-of-experts design for better quality at the same compute, adds a hybrid TI2V-5B model that runs on a single 4090, and includes Animate for character animation and subject replacement. Higher steps and a moderate guidance scale give cleaner motion, and the workflows expose those settings so you are not guessing.

Everything here is open and editable. Run a Wan workflow on Comfy Cloud with nothing to install when you do not have the GPU for it, or download the graph and run it locally for full control. Either way you can swap models, tune the parameters, and re-run until the clip is right, with no watermark on the result.

InfiniteTalk: Audio-Driven Full-Body Video Dubbing
Wan2.1 InfiniteTalk

Wan2.1 InfiniteTalk

Open source

Open source under Apache 2.0, so you can run it locally or on Comfy Cloud.

Wan2.7: Text to Video
WAN

Wan2.7

Video and lip-sync

Text-to-video, image-to-video, and speech-to-video with lip-sync (S2V).

Reproducible, tunable, yours

Every step is exposed and adjustable, and the exact graph is yours to reuse across projects.

Start free. Upgrade when you’re ready.

See pricing plans

Capabilities

What You Can Create

Wan is Alibaba Tongyi Lab's open-source (Apache 2.0) video generation family, covering text-to-video, image-to-video, and speech-to-video across the 2.1 and 2.2 releases.

Wan 2.1 sizes

Wan 2.1: 1.3B size on ~8GB VRAM, 14B on 24GB+, at 480p and 720p.

Try workflow
Wan2.1 VACE Text to Video
Wan2.1

Wan 2.2 MoE

Wan 2.2: mixture-of-experts for higher quality, plus a TI2V-5B that runs on a single 4090.

Try workflow
Wan-Move Motion-Control Image to Video
Wan

Fun control variants

Fun variants add camera control and depth, pose, or canny conditioning.

Try workflow
Wan2.1 SCAIL
Wan2.1 SCAIL
HOWComfyWORKS

Connect models, processing steps, and outputs on a canvas where every decision is visible and every step is inspectable.

Start from a community template or build from scratch.

Comparison

Why ComfyUI Workflows

  • Generation, animation, and editing in one family

    Text-to-video, image-to-video, speech-to-video with lip-sync, and VACE editing all run from the same Wan graphs.

  • Pick your variant and settings

    ONLY ONComfyCLOUD

    Switch between Wan 2.1 1.3B, the 14B sizes, and the Wan 2.2 TI2V-5B hybrid, then tune frames, steps, and guidance scale.

  • Open weights under Apache 2.0

    Wan is released under Apache 2.0, so you can run it locally on your own GPU or on Comfy Cloud and use the output commercially.

  • No watermark on your clips

    Wan generates at 24 frames per second across multiple aspect ratios, and the clips you export ship with no watermark.

Applications

What people use it for

What people generate with Wan, the open Apache 2.0 video family.

Wan 2.1 Text to Video

Text to video

Wan2.1

Generate short wan video clips from a written prompt with wan text to video

Frequently Asked Questions

Wan FAQ

Ready to create?

Start generating with this workflow in seconds

Try Comfy Cloud