ComfyUI: Generating a 52-second avatar video

Demo summary
The user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.
Step-by-step
- Input your script into the text-to-speech section or use the Ollama inputs for generated content
- Select an image for the avatar
- Set the video dimensions
- Select a reference audio sample to clone the voice
- Generate the audio from the text
- Initiate the sampling step to generate the video
Options
- Use Ollama to generate content via 'speaker name' and 'topic' inputs
- Manually type a script for more control
- Set dimensions to 720p if your setup supports it
Watch out for
- A reference voice audio sample is required to clone the voice
Tips
- Use 480p dimensions to ensure the workflow runs smoothly on most setups
- Follow the numbered labels (Step 1, Step 2) in the workflow to stay organized
Highlights
“480p should work for most setups”
All demos from “InfiniteTalk in ComfyUI Tutorial – The Next Level of AI Talking Avatar!”
5:151:43Configuring InfiniteTalk nodes in ComfyUIThe creator demonstrates how to use the WAN Video Wrapper in ComfyUI, selecting the Infinite Talk model loader and configuring the single-person talking avatar settings.ComfyUI· AI Animation Generator
7:081:19Automating podcast videos with Ollama and Chatterbox SRTA workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.ComfyUI· AI Avatar Video Generator
8:431:16Generating a 52-second avatar videoCurrentThe user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.ComfyUI· AI Avatar Video Generator- Watch “InfiniteTalk in ComfyUI Tutorial – The Next Level of AI Talking Avatar!” →
AI Avatar Video Generator
25:063:04Create talking avatars with InfiniteTalkThe demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.pixaroma
5:421:22Configure InfiniteTalk for talking-head generationThe user demonstrates setting up the InfiniteTalk model loader, uploading a source image and audio file, and calculating the required frame count for a 25-second clip.MDMZ
7:321:10Animate AI avatar with InfiniteTalk settingsThe creator adjusts the audio scale for expressiveness, adds a descriptive prompt for movement style, and configures the One Video Sampler steps before running the final generation.MDMZ
0:441:10Generate a talking head video with Infinite Talk in ComfyUIThe creator demonstrates the basic Infinite Talk workflow by uploading a portrait image and an audio file, adjusting the prompt and resolution, and rendering a lip-synced video.pixaroma
7:081:19Automating podcast videos with Ollama and Chatterbox SRTA workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.Benji’s AI Playground
8:431:16Generating a 52-second avatar videoCurrentThe user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.Benji’s AI Playground
7:580:38Generating the final talking videoThe creator shows the KSampler settings using Lightning LoRA for fast 7-step generation and displays the final 53-second talking head result.Aiconomist
ComfyUI