Text-to-Video generation with Hunyuan Video 1.5 in ComfyUI

Demo summary
The creator demonstrates a ComfyUI workflow using Hunyuan Video 1.5 (FP16) and Qwen2.5-VL for prompt generation to create a 720p video of a horse and rider.
Step-by-step
- Load the Hunyuan Video 1.5 main model, selecting FP16 for quality or FP8 if you have low VRAM.
- Enable Sage Attention to increase generation speed and quality.
- Load the dual text encoders, specifically selecting the Qwen2.5-VL and BYT5 models.
- Set the model type to 'hunyuan VIDEO 15' in the encoder node.
- Configure the latent node resolution to 1280x720 with a frame length of 33.
- Set the CFG to 6.0 and the Shift value to 9 for 720p generation.
- Configure the KSampler with 20 steps, the Euler sampler, and the Simple scheduler.
Options
- Use FP8 instead of FP16 to avoid out-of-memory errors.
- Use 50 sampling steps for higher quality if time is not a concern.
- Write prompts manually or use the Qwen VL model to generate them from an image.
Watch out for
- You must select 'hunyuan VIDEO 15' in the text encoder type or it will not work correctly.
- The Shift value must be adjusted according to the specific resolution list (e.g., 9 for 720p).
Tips
- Enable Sage Attention for a faster and higher-quality result.
- Use 20 sampling steps instead of 50 to save significant time while maintaining good results.
- Refer to the official GitHub guide for the specific prompt engineering format.
- Use a prompt engineering workflow with Qwen VL for better descriptive prompts.
Highlights
“we can then achieve a fantastic effect”
All demos from “极简工作流!腾讯混元 1.5 (Hunyuan Video) 首发测评:原生 1080P + 显存优化全攻略”
5:282:26Text-to-Video generation with Hunyuan Video 1.5 in ComfyUICurrentThe creator demonstrates a ComfyUI workflow using Hunyuan Video 1.5 (FP16) and Qwen2.5-VL for prompt generation to create a 720p video of a horse and rider.ComfyUI· AI Animation Generator
7:542:071080P Latent Upscaling with Hunyuan SR ModelThe video shows how to use the Hunyuan Video Latent Upscale node and the SR (Super Resolution) model to enhance a 720p video to 1080p using a dual-sampler noise injection technique.ComfyUI· AI Video Upscaler
10:012:01Image-to-Video generation using Hunyuan 1.5 DistilledThe user demonstrates an I2V workflow in ComfyUI using the Hunyuan 1.5 distilled model, Google's Clip Vision for image encoding, and the 'image to video' node to animate a portrait.ComfyUI· Image to Video- Watch “极简工作流!腾讯混元 1.5 (Hunyuan Video) 首发测评:原生 1080P + 显存优化全攻略” →
AI Animation Generator
1:420:59Seed hunting with a multi-stage LTX 2.3 workflowThe creator demonstrates his custom ComfyUI workflow that generates four low-resolution LTX 2.3 samples simultaneously to find a 'golden seed' before upscaling to 1080p.Fox•Fur•Essence Films
16:361:56Setting up HunyuanVideo 1.5 in ComfyUIThe video demonstrates how to update ComfyUI and import the HunyuanVideo 1.5 JSON workflow files to create a node-based generation environment.AI Search
20:151:53Text-to-Video generation in ComfyUIA step-by-step demo of configuring the Hunyuan nodes in ComfyUI, entering a prompt for a 'giant cat', and rendering the final 720p video.AI Search
29:151:33Running HunyuanVideo with GGUF (Low VRAM)The video shows how to use the GGUF loader node to run a compressed version of HunyuanVideo 1.5, enabling video generation on GPUs with as little as 6GB of VRAM.AI Search
0:510:36Configure Infinite Talk and Wan 2.1 models in ComfyUIThe user demonstrates loading the Infinite Talk model alongside the Wan 2.1 I2V 14B model within ComfyUI, including enabling block swap and torch compile for VRAM optimization.Olares
1:551:13Configure sampling and window settings for long video generationThe user walks through the Wan Video Wrapper sampling node, explaining how to set frame window size, motion frame overlap, and start steps for consistent video generation.Olares
0:310:46Generate cinematic video with LTX MSR workflowThe creator demonstrates using the LTX MSR workflow in ComfyUI to generate a video from multiple reference images and a prompt, highlighting the 3D camera movement and character consistency.Apex Artist
0:290:24Load source footage and models in ComfyUIThe user demonstrates importing source video footage and loading the necessary model nodes including WAN Video, VAE, and Clip Vision within the ComfyUI interface.ComfyUI
0:431:12Configure LTX-2.3 MSR LoRA and Prompt Relay in ComfyUIThe creator demonstrates setting up a ComfyUI workflow using the LTX-2.3 model with the MSR LoRA and Prompt Relay nodes to manage video generation on 8GB of VRAM.bigboss97
4:261:28Overview of the TensNodes Consistent Character Workflow in ComfyUIThe creator walks through a four-stage ComfyUI workflow utilizing TensNodes to process reference images and text for stable LTX video generation.SOTAI
3:411:25Configure LTX 2.3 foundational parametersA walkthrough of setting dimensions, frame counts, and loading core models including the LTX 2.3 distill model and audio VAE within a ComfyUI workflow.SOTAI
5:060:37Process visual and audio inputs in ComfyUIThe demo shows how to use the LTX V image-to-video condition node and empty latent audio node to prepare data for the generation engine.SOTAI
ComfyUI