demobook

ComfyUI: Create talking avatars with InfiniteTalk

Demo summary

The demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.

Step-by-step

  1. Download the Wan 2.1 model and save it in the models/diffusion_models/wan2.1 folder.
  2. Download the InfiniteTalk FP16 model and save it in the models/model_patches folder.
  3. Download the audio encoder and save it in the models/audio_encoders folder.
  4. Ensure the text encoder and VAE are in their respective folders, then press R to refresh node definitions.
  5. Load an audio file into the workflow.
  6. Upload a portrait image where the mouth is clearly visible.
  7. Set the video size to HD and configure the prompt for character talking.
  8. Run the workflow to generate and extend the video to the desired duration.

Options

  • Use 'start time' to begin the audio at a specific second instead of zero
  • Enable 'trim to audio' to cut the video duration to match the audio file length
  • Enable saving of intermediate 3-second video segments instead of just the final output

Watch out for

  • The workflow can only generate HD videos.
  • The prompt cannot handle complex animations like walking or dancing while talking.
  • Longer videos require adding more nodes as each node extends the video by 3 seconds.

Tips

  • Use portrait images where the mouth is visible for the best results.
  • The Chinese audio encoder model is the correct one to use even for English audio.
  • Organize models into specific subfolders (e.g., wan2.1) to ensure the workflow recognizes them correctly.

All demos from “ComfyUI Video Models: InfiniteTalk + Wan 2.2 + SCAIL + LTX-2 (Ep06)

  1. 5:244:57Image-to-video generation with Wan 2.2The user demonstrates how to use the Wan 2.2 GGUF model in ComfyUI to animate a static image of a woman based on a detailed text prompt.ComfyUIImage to Video
  2. 13:071:21Applying LoRA to Wan 2.2 video generationThe video shows how to integrate a LoRA (Low-Rank Adaptation) into a Wan 2.2 workflow to achieve specific cinematic movements like a face zoom.ComfyUIAI Animation Generator
  3. 15:344:59Character replacement with Wan AnimateThe creator demonstrates using Wan Animate and SAM 3 to mask a character in a reference video and replace them with a new character from a static image while maintaining the original motion.ComfyUIVideo to Video
  4. 21:072:43Cartoon animation with Wan SCAILThe user shows how to animate a cartoon ballerina using a real-person video as a motion reference via the Wan SCAIL workflow in ComfyUI.ComfyUIVideo to Video
  5. 25:063:04Create talking avatars with InfiniteTalkCurrentThe demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.ComfyUIAI Avatar Video Generator
  6. 31:133:44Text-to-video with LTX-2The video walks through setting up the LTX-2 model in ComfyUI to generate high-resolution video clips from text prompts and images.ComfyUIAI Animation Generator
  7. 36:000:39Singing characters with LTX-2 and custom audioThe creator demonstrates how to use LTX-2 to make a character sing by providing a custom audio file and a specific singing prompt.ComfyUIAI Lip Sync Generator
  8. 37:281:08Upscale video with Seed-V2The user demonstrates upscaling a low-resolution AI-generated video to Full HD using the Seed-V2 workflow in ComfyUI for improved sharpness.ComfyUIAI Video Upscaler
  9. 39:075:27Cloud-based ComfyUI on RunPod/RunHubThe video shows how to run complex video workflows in the cloud using RunHub AI, demonstrating the interface and execution of InfiniteTalk and Wan 2.2 without local hardware.ComfyUIAI Animation Generator
  10. 44:340:52Frame interpolation for smoother motionThe user demonstrates a workflow to double the frame rate of a 16fps video to 32fps to create smoother motion in AI-generated clips.ComfyUIAI Video Interpolation
  11. Watch “ComfyUI Video Models: InfiniteTalk + Wan 2.2 + SCAIL + LTX-2 (Ep06)” →

AI Avatar Video Generator

  1. 25:063:04Create talking avatars with InfiniteTalkCurrentThe demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.pixaroma
  2. 5:421:22Configure InfiniteTalk for talking-head generationThe user demonstrates setting up the InfiniteTalk model loader, uploading a source image and audio file, and calculating the required frame count for a 25-second clip.MDMZ
  3. 7:321:10Animate AI avatar with InfiniteTalk settingsThe creator adjusts the audio scale for expressiveness, adds a descriptive prompt for movement style, and configures the One Video Sampler steps before running the final generation.MDMZ
  4. 0:441:10Generate a talking head video with Infinite Talk in ComfyUIThe creator demonstrates the basic Infinite Talk workflow by uploading a portrait image and an audio file, adjusting the prompt and resolution, and rendering a lip-synced video.pixaroma
  5. 7:081:19Automating podcast videos with Ollama and Chatterbox SRTA workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.Benji’s AI Playground
  6. 8:431:16Generating a 52-second avatar videoThe user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.Benji’s AI Playground
  7. 7:580:38Generating the final talking videoThe creator shows the KSampler settings using Lightning LoRA for fast 7-step generation and displays the final 53-second talking head result.Aiconomist