demobook

ComfyUI: Automating podcast videos with Ollama and Chatterbox SRT

Demo summary

A workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.

Step-by-step

  1. Generate content using the Ollama node or type manual text into the input field
  2. Use the bypass group switch to select between the LLM output or manual text
  3. Pass the text to the Chatterbox SRT node to generate audio
  4. Allow the workflow to dynamically calculate video length and FPS based on the audio duration
  5. Set the sampling step to 4
  6. Run the sampler to generate the talking avatar video

Options

  • Generate content via Ollama
  • Input manual text
  • Use frame interpolation to double the FPS and reduce flickering

Watch out for

  • The original Chatterbox project is currently broken and incompatible with the latest ComfyUI updates
  • The workflow may produce minor issues like fast blinking or eye flickering in the raw output

Tips

  • Use the Chatterbox SRT custom node pack from GitHub as a working alternative
  • Apply frame interpolation to fix eye flickering and significantly improve video quality

All demos from “InfiniteTalk in ComfyUI Tutorial – The Next Level of AI Talking Avatar!

  1. 5:151:43Configuring InfiniteTalk nodes in ComfyUIThe creator demonstrates how to use the WAN Video Wrapper in ComfyUI, selecting the Infinite Talk model loader and configuring the single-person talking avatar settings.ComfyUIAI Animation Generator
  2. 7:081:19Automating podcast videos with Ollama and Chatterbox SRTCurrentA workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.ComfyUIAI Avatar Video Generator
  3. 8:431:16Generating a 52-second avatar videoThe user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.ComfyUIAI Avatar Video Generator
  4. Watch “InfiniteTalk in ComfyUI Tutorial – The Next Level of AI Talking Avatar!” →

AI Avatar Video Generator

  1. 25:063:04Create talking avatars with InfiniteTalkThe demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.pixaroma
  2. 5:421:22Configure InfiniteTalk for talking-head generationThe user demonstrates setting up the InfiniteTalk model loader, uploading a source image and audio file, and calculating the required frame count for a 25-second clip.MDMZ
  3. 7:321:10Animate AI avatar with InfiniteTalk settingsThe creator adjusts the audio scale for expressiveness, adds a descriptive prompt for movement style, and configures the One Video Sampler steps before running the final generation.MDMZ
  4. 0:441:10Generate a talking head video with Infinite Talk in ComfyUIThe creator demonstrates the basic Infinite Talk workflow by uploading a portrait image and an audio file, adjusting the prompt and resolution, and rendering a lip-synced video.pixaroma
  5. 7:081:19Automating podcast videos with Ollama and Chatterbox SRTCurrentA workflow demonstration showing how to use Ollama for text generation and Chatterbox SRT for text-to-speech to automatically drive the InfiniteTalk avatar length and audio sync.Benji’s AI Playground
  6. 8:431:16Generating a 52-second avatar videoThe user walks through the step-by-step process of inputting a script, selecting a reference voice for cloning, and initiating the sampling process to generate a nearly one-minute video.Benji’s AI Playground
  7. 7:580:38Generating the final talking videoThe creator shows the KSampler settings using Lightning LoRA for fast 7-step generation and displays the final 53-second talking head result.Aiconomist