demobook

ComfyUI: Generate character-consistent video with Wan 2.6 R2V

Demo summary

The creator demonstrates a ComfyUI workflow using the Wan 2.6 API node to transform a 5-second reference video of himself into a new scene while maintaining character consistency and native audio lip-sync.

Step-by-step

  1. Open the ComfyUI workflow containing the Load Video, Wan 2.6 API, and Video Combine nodes.
  2. Upload a short reference video of your subject into the Load Video node.
  3. Enter a prompt describing the new scene, setting, action, and mood.
  4. Specify the dialogue in the prompt if you want the character to say specific words.
  5. Run the workflow to generate the 5 or 10-second clip.

Options

  • Input up to three reference videos to include multiple subjects or specific objects
  • Let the model generate random speech or provide a specific script for lip-syncing
  • Reference people, animals, characters, or objects

Watch out for

  • Using the Wan 2.6 API node incurs a cost (approximately 50 cents per video)
  • Reference videos are limited to 5 or 10-second clips at up to 1080p resolution

Tips

  • Cram as much information as possible into the reference video by capturing different angles and expressions.
  • Move the camera around the subject (like a 360° walk-around) to help the model understand the subject's form and movement.
  • Include talking in the reference video to enable the native audio lip-sync and voice replication features.

Highlights

So we have a very simple workflow to work with in Comfy here... this should be fairly simple, right?

All demos from “Wan 2.6 is HERE! R2V is AWESOME!

  1. 1:003:01Generate character-consistent video with Wan 2.6 R2VCurrentThe creator demonstrates a ComfyUI workflow using the Wan 2.6 API node to transform a 5-second reference video of himself into a new scene while maintaining character consistency and native audio lip-sync.ComfyUIVideo to Video
  2. 5:170:33Multi-reference video generation in Wan 2.6A demonstration of using multiple input videos (a subject, a motorcycle, and a location) within the Wan 2.6 workflow to generate a complex scene of the subject riding a specific bike in a specific setting.WWanVideo to Video
  3. Watch “Wan 2.6 is HERE! R2V is AWESOME!” →

Video to Video

  1. 8:392:29Comparing LTX 2.3 Vanilla vs. TensNodes for Facial ExpressionsA side-by-side demonstration shows the TensNodes workflow maintaining a consistent facial structure during a smile and wave, whereas the vanilla LTX model drifts from the reference image.SOTAI
  2. 15:344:59Character replacement with Wan AnimateThe creator demonstrates using Wan Animate and SAM 3 to mask a character in a reference video and replace them with a new character from a static image while maintaining the original motion.pixaroma
  3. 21:072:43Cartoon animation with Wan SCAILThe user shows how to animate a cartoon ballerina using a real-person video as a motion reference via the Wan SCAIL workflow in ComfyUI.pixaroma
  4. 4:031:00Generating Guide Videos with BFS NodeThe demo shows how to use the BFS Node extension to combine a reference face image and a driving video into a green-screen guide video required for the LTX sampling process.Veteran AI
  5. 5:030:47Injecting Latent Constraints with Add Guide MultiThe video demonstrates using the 'Add Guide Multi' node from the KJ Node extension to inject the guide video constraints and VAE-encoded latents into the sampling pipeline.Veteran AI
  6. 15:352:08Final video generation with Wan 2.1The demonstration shows how to use the Wan 2.1 V2V model to combine the generated 3D motion and a reference character image into a final, high-quality AI video animation.Benji’s AI Playground
  7. 8:521:04Replacing a dancing character with MochaA concrete demonstration of replacing a woman dancing in a video with a new character using the Mocha workflow in ComfyUI.SOTAI
  8. 12:450:58Replacing a human with an animated werewolfThe video demonstrates replacing a human actor with a stylized werewolf character while maintaining consistent textures and movement.SOTAI
  9. 1:003:01Generate character-consistent video with Wan 2.6 R2VCurrentThe creator demonstrates a ComfyUI workflow using the Wan 2.6 API node to transform a 5-second reference video of himself into a new scene while maintaining character consistency and native audio lip-sync.Sebastian Kamph
  10. 1:551:00Adding reference images for character consistencyThe video shows how to use the MSR custom node to input up to four reference images and a background image to maintain character facial consistency across video frames.bigboss97
  11. 2:280:44Swap human character for 3D Pixar model in MochaThe creator demonstrates swapping a real person in a video with a 3D Pixar-style character using Mocha, showing the transfer of lip movements and hand gestures.AI Search
  12. 20:110:27Generating final character swap videoThe video shows the K-Sampler process and the final rendered output of a character replacement task within ComfyUI.AI Search