demobook

ComfyUI: Text-to-image generation with Z-Turbo and Qwen

Demo summary

Demonstration of the Z-Turbo model and Qwen text encoder to generate specific advertising text on the side of a 3D tram model within a ComfyUI workflow.

Step-by-step

  1. Load the Z-Turbo model into the workflow
  2. Set the text encoder to the Qwen model
  3. Enter a detailed prompt specifying the text to appear on the image
  4. Connect a ControlNet node using a depth model
  5. Draw the input image or screenshot in the workspace
  6. Adjust the denoising strength for the rendering
  7. Press Control+B to bypass or turn off the LoRA node if you are not using one

Options

  • Use Depth Anything if Lotus Depth is not working
  • Add a preview node to see the drawing input
  • Toggle the LoRA node on or off depending on your model setup

Watch out for

  • Spelling mistakes in generated text are to be expected
  • The background may appear flat or low-detail if the input screenshot lacks background information

Tips

  • Use the Qwen text encoder for better prompt coherence and more specific, detailed prompts
  • Set the ControlNet strength to approximately 0.65

Highlights

it's kind of amazing that it can do any of this

All demos from “Local "AI Rendering" Using 3D - Explained by a Human

  1. 3:041:18Enhance 3D clay models with Flux and ControlNetThe creator demonstrates using the Flux model in ComfyUI with a Depth ControlNet to transform a gray 3D massing model into a detailed streetcar scene while maintaining structural integrity.ComfyUIImage to Image
  2. 6:401:29Train a custom LoRA on CivitaiA step-by-step walkthrough of uploading a dataset of streetcar images to Civitai, using auto-labeling, and configuring training settings to create a custom LoRA model.CivitaiAI Image Generator
  3. 8:512:02Image-to-image enhancement with denoisingThe video shows how to use the 'denoising' strength in ComfyUI to blend a base 3D rendering with AI-generated details like autumn trees and realistic people.ComfyUIImage to Image
  4. 12:211:24Text-to-image generation with Z-Turbo and QwenCurrentDemonstration of the Z-Turbo model and Qwen text encoder to generate specific advertising text on the side of a 3D tram model within a ComfyUI workflow.ComfyUIText to Image
  5. 15:063:12Transform 3D scenes with Qwen2-VL (Qwen-Edit)The creator uses the Qwen-Edit workflow to perform complex scene modifications, such as changing weather to rain or snow, while preserving the original 3D geometry and text.ComfyUIAI Inpainting
  6. 20:221:07Fix text and faces using Crop and StitchA demonstration of a custom 'crop image' node to isolate specific areas like signs or faces, regenerate them at native resolution, and stitch them back into the high-res image.ComfyUIAI Inpainting
  7. 23:511:57Animate 3D renders with Wan 2.1The creator walks through a Wan 2.1 video generation workflow in ComfyUI, using the Painterly I2V node to control motion speed and animate a static 3D render.ComfyUIImage to Video
  8. Watch “Local "AI Rendering" Using 3D - Explained by a Human” →

Text to Image

  1. 2:501:21Generate an AI image from a text promptThe video shows the process of entering positive and negative prompts, configuring image dimensions, and clicking 'Run' to generate a chocolate chip cookie image in ComfyUI.Kevin Stratvert
  2. 12:211:24Text-to-image generation with Z-Turbo and QwenCurrentDemonstration of the Z-Turbo model and Qwen text encoder to generate specific advertising text on the side of a 3D tram model within a ComfyUI workflow.Matt Hallett Visual
  3. 20:151:27Text-to-image generation with PiD DIT modelThe creator demonstrates the standalone PiD text-to-image workflow, generating an image of a leopard in a jungle directly from a prompt using the lightweight DIT model.AI Search
  4. 22:350:32Generating an image with Flux modelDemonstrates a text-to-image generation using the Flux model in ComfyUI, including prompt entry and resolving connection errors.Sebastian Kamph
  5. 7:390:44Basic text-to-image generation in ComfyUIThe user demonstrates a basic text-to-image workflow using the RealVisXL model to generate an image of a castle in a forest.AI Search
  6. 7:020:23Text-to-image generation with FluxThe user demonstrates Flux's prompt adherence by generating an image of an old TV with the word 'flux' on it in an abandoned workshop.Artificial Images
  7. 4:250:43Generate images with Stable Diffusion 3.5The video demonstrates entering a text prompt, selecting CLIP models, and clicking 'Queue' to generate a high-definition image using Stable Diffusion 3.5.AIPURE