ComfyUI: Creating a character scene with DW Pose and Qwen2-VL

Demo summary
The creator shows how to extract a pose from the rendered motion using DW Pose and generate a descriptive caption using Qwen2-VL to prepare for final image generation.
Step-by-step
- Load the reference image via URL, local file path, or the ComfyUI input folder
- Connect the image to an Image Switch node to toggle between references
- Run the image through the Qwen2-VL node to autogenerate a text description
- Disable the Quant Image Edit model group to extract the caption first
- Re-enable the Quant Image Edit group using the latest 200511 model
- Input the DW pose map into the second image slot
- Enter the desired visual style and action scene details into the text box
Options
- Load images from a URL
- Load images from a local file path
- Load images from the ComfyUI input folder
- Toggle between different references using an Image Switch node
Watch out for
- The Quant Image Edit model group must be disabled initially if you only want the caption without running a full edit
Tips
- Copy the autogenerated description from Qwen2-VL to use as parts of your final prompts later
- Use an Image Switch node for convenience when working with multiple reference images
Highlights
“This is super helpful because I can copy that description and use parts of it in my final prompts later.”
All demos from “Hunyuan Motion 1.0 - This Free And Open Source AI Creates 3D Animation In Seconds!”
6:471:07Setting up Hunyuan Motion nodes in ComfyUIThe creator demonstrates how to add the 'load model' and 'motion generate' nodes in ComfyUI, connect them to a 3D preview node, and view a real-time skeleton animation of the generated motion.ComfyUI· AI Animation Generator
8:210:59Generating 3D motion from text promptsUsing the Hunyuan Motion 1.0 model in ComfyUI, a text prompt describing a person jumping between buildings is used to generate a specific 3D motion sequence.ComfyUI· AI Animation Generator
13:101:04Creating a character scene with DW Pose and Qwen2-VLCurrentThe creator shows how to extract a pose from the rendered motion using DW Pose and generate a descriptive caption using Qwen2-VL to prepare for final image generation.ComfyUI· AI Image Generator
15:352:08Final video generation with Wan 2.1The demonstration shows how to use the Wan 2.1 V2V model to combine the generated 3D motion and a reference character image into a final, high-quality AI video animation.ComfyUI· Video to Video- Watch “Hunyuan Motion 1.0 - This Free And Open Source AI Creates 3D Animation In Seconds!” →
AI Image Generator
1:560:39Load a text-to-image workflow and download missing modelsThe user demonstrates how to load a default text-to-image template in ComfyUI and use the 'show missing models' feature to download required AI checkpoints.Kevin Stratvert
4:110:29Manage generated assets in ComfyUIThe presenter demonstrates how to access the assets panel using the 'A' shortcut to view, save, or delete previously generated images.Kevin Stratvert
15:005:45Generate image with Hermes Agent and ComfyUIThe host uses the Hermes Agent TUI to request a 1024x1024 image of a tree using the Zimage Turbo model, which the agent executes by generating and running a ComfyUI workflow.ComfyUI
6:420:59Configure Gemma 2B Text Encoder in ComfyUIThe creator shows how to load the Gemma 2B text encoder model into the ComfyUI workflow and select it from the node dropdown menu.AI Search
4:270:30Copying and replacing prompts in ComfyUIThe user demonstrates a custom node feature that allows hovering over prompt examples to copy and then replace the current prompt in the workflow with a single click.pixaroma
13:030:51Managing node colors with PixaRoma nodesThe video demonstrates UI updates for PixaRoma nodes in ComfyUI, including selecting color swatches, copying/pasting colors between nodes, and setting favorite colors.pixaroma
4:150:35Launch ComfyUI on WindowsThe creator demonstrates how to launch the portable version of ComfyUI by running the 'run_nvidia_gpu' batch file and accessing the interface via a web browser.WINBUSH
10:071:30Browse and load ComfyUI templatesThe creator demonstrates how to navigate the templates menu, filter for local models, and load a Flux 2 image generation workflow.WINBUSH
0:000:33Configuring MSR mode in LTX DirectorThe user demonstrates setting up the Director node in ComfyUI for MSR (Multiple Subject Reference) mode, including bypassing clean latent slices and enabling the MSR LoRA.RedShade
5:090:19Loading LTX Director V5 workflow templateThe video shows how to search for and load the LTX Director V5 template within ComfyUI and input prompts for character generation.RedShade
4:251:21Iterative rerolling at specific generation stagesThe video shows how to use the 'reroll' button at specific stages of the workflow to change motion or visual quality without re-generating the initial seed samples.Fox•Fur•Essence Films
9:173:04Install ComfyUI using Easy InstallThe instructor demonstrates how to download the ComfyUI Easy Install zip from GitHub, extract it, and run the BAT file to set up a portable local installation.pixaroma
ComfyUI