ComfyUI: Text-to-video with LTX-2

Demo summary
The video walks through setting up the LTX-2 model in ComfyUI to generate high-resolution video clips from text prompts and images.
Step-by-step
- Download the LTX-2 model and place it in a subfolder named 'LTX2' within the 'checkpoints' folder.
- Save the text encoder model into the 'text_encoders' folder.
- Save the latent upscale model into the 'latent_upscale_models' folder.
- Press 'Refresh' or 'R' in ComfyUI to update the node definitions.
- Load the image using the upload button and adjust the video dimensions to match the image ratio.
- Set the total frame count to a number divisible by 8, then add 1 extra frame (e.g., for 25 fps).
- Enter a descriptive prompt for the desired animation.
- Run the workflow to process the two-stage sampling and upscaling.
Options
- Generate video with audio (note: may sound robotic)
- Generate up to Full HD resolution if VRAM allows
Watch out for
- The LTX-2 model must be in the 'checkpoints' folder; it will not be detected if placed in 'diffusion_models'.
- All video dimensions (width and height) must be multiples of 32.
- Prompt understanding is not as strong as the Juan model and it may make more mistakes due to faster generation.
Tips
- Use the provided frame count list for durations up to 15 seconds to ensure compatibility.
- Consult external prompting guides specifically for LTX to improve results.
- Avoid manually tweaking the sampler settings if the default workflow is already producing functional results.
- Stick to standard resolutions as Full HD results may not be high quality despite VRAM availability.
Highlights
“That happens when you work hours in Comfy UI with all these complex models. You always forget to change something.”
All demos from “ComfyUI Video Models: InfiniteTalk + Wan 2.2 + SCAIL + LTX-2 (Ep06)”
5:244:57Image-to-video generation with Wan 2.2The user demonstrates how to use the Wan 2.2 GGUF model in ComfyUI to animate a static image of a woman based on a detailed text prompt.ComfyUI· Image to Video
13:071:21Applying LoRA to Wan 2.2 video generationThe video shows how to integrate a LoRA (Low-Rank Adaptation) into a Wan 2.2 workflow to achieve specific cinematic movements like a face zoom.ComfyUI· AI Animation Generator
15:344:59Character replacement with Wan AnimateThe creator demonstrates using Wan Animate and SAM 3 to mask a character in a reference video and replace them with a new character from a static image while maintaining the original motion.ComfyUI· Video to Video
21:072:43Cartoon animation with Wan SCAILThe user shows how to animate a cartoon ballerina using a real-person video as a motion reference via the Wan SCAIL workflow in ComfyUI.ComfyUI· Video to Video
25:063:04Create talking avatars with InfiniteTalkThe demonstration shows how to sync a static character image with an audio file to create a talking avatar using the InfiniteTalk model and Wan 2.1.ComfyUI· AI Avatar Video Generator
31:133:44Text-to-video with LTX-2CurrentThe video walks through setting up the LTX-2 model in ComfyUI to generate high-resolution video clips from text prompts and images.ComfyUI· AI Animation Generator
36:000:39Singing characters with LTX-2 and custom audioThe creator demonstrates how to use LTX-2 to make a character sing by providing a custom audio file and a specific singing prompt.ComfyUI· AI Lip Sync Generator
37:281:08Upscale video with Seed-V2The user demonstrates upscaling a low-resolution AI-generated video to Full HD using the Seed-V2 workflow in ComfyUI for improved sharpness.ComfyUI· AI Video Upscaler
39:075:27Cloud-based ComfyUI on RunPod/RunHubThe video shows how to run complex video workflows in the cloud using RunHub AI, demonstrating the interface and execution of InfiniteTalk and Wan 2.2 without local hardware.ComfyUI· AI Animation Generator
44:340:52Frame interpolation for smoother motionThe user demonstrates a workflow to double the frame rate of a 16fps video to 32fps to create smoother motion in AI-generated clips.ComfyUI· AI Video Interpolation- Watch “ComfyUI Video Models: InfiniteTalk + Wan 2.2 + SCAIL + LTX-2 (Ep06)” →
AI Animation Generator
1:420:59Seed hunting with a multi-stage LTX 2.3 workflowThe creator demonstrates his custom ComfyUI workflow that generates four low-resolution LTX 2.3 samples simultaneously to find a 'golden seed' before upscaling to 1080p.Fox•Fur•Essence Films
16:361:56Setting up HunyuanVideo 1.5 in ComfyUIThe video demonstrates how to update ComfyUI and import the HunyuanVideo 1.5 JSON workflow files to create a node-based generation environment.AI Search
20:151:53Text-to-Video generation in ComfyUIA step-by-step demo of configuring the Hunyuan nodes in ComfyUI, entering a prompt for a 'giant cat', and rendering the final 720p video.AI Search
29:151:33Running HunyuanVideo with GGUF (Low VRAM)The video shows how to use the GGUF loader node to run a compressed version of HunyuanVideo 1.5, enabling video generation on GPUs with as little as 6GB of VRAM.AI Search
0:510:36Configure Infinite Talk and Wan 2.1 models in ComfyUIThe user demonstrates loading the Infinite Talk model alongside the Wan 2.1 I2V 14B model within ComfyUI, including enabling block swap and torch compile for VRAM optimization.Olares
1:551:13Configure sampling and window settings for long video generationThe user walks through the Wan Video Wrapper sampling node, explaining how to set frame window size, motion frame overlap, and start steps for consistent video generation.Olares
0:310:46Generate cinematic video with LTX MSR workflowThe creator demonstrates using the LTX MSR workflow in ComfyUI to generate a video from multiple reference images and a prompt, highlighting the 3D camera movement and character consistency.Apex Artist
0:290:24Load source footage and models in ComfyUIThe user demonstrates importing source video footage and loading the necessary model nodes including WAN Video, VAE, and Clip Vision within the ComfyUI interface.ComfyUI
0:431:12Configure LTX-2.3 MSR LoRA and Prompt Relay in ComfyUIThe creator demonstrates setting up a ComfyUI workflow using the LTX-2.3 model with the MSR LoRA and Prompt Relay nodes to manage video generation on 8GB of VRAM.bigboss97
4:261:28Overview of the TensNodes Consistent Character Workflow in ComfyUIThe creator walks through a four-stage ComfyUI workflow utilizing TensNodes to process reference images and text for stable LTX video generation.SOTAI
3:411:25Configure LTX 2.3 foundational parametersA walkthrough of setting dimensions, frame counts, and loading core models including the LTX 2.3 distill model and audio VAE within a ComfyUI workflow.SOTAI
5:060:37Process visual and audio inputs in ComfyUIThe demo shows how to use the LTX V image-to-video condition node and empty latent audio node to prepare data for the generation engine.SOTAI
ComfyUI