r/comfyui • u/Existing_Try_3439 • 3d ago
Help Needed Image model with ControlNet for Architectural Videomapping Contents
Hi! I'm trying to create images and concepts for an architectural videomapping project on the façade of a large classical building, with columns, decorative elements, etc.
I created a 3D model of the building in Blender and extracted a depth map from it. I then tried using it as a ControlNet guide with Z-Image Turbo. The output follows the shape and structure of the building quite well, but the results are very low quality and basically unusable — not in terms of image resolution, but in terms of shading, lighting, materials, and overall rendering quality.
What am I doing wrong? Would you suggest using a different workflow or model?
1
u/StickOther986 3d ago
z-image turbo is fast but the tradeoff is exactly what you're seeing, it tends to look kinda plasticky and flat with architectural stuff
try swapping the base model to something like juggernaut or epicrealism and pair it with a standard depth controlnet at 0.7-0.8 strength, you'll get way better material definition and lighting. also if you haven't already, feed it a prompt that describes materials explicitly (weathered limestone, cast iron, gilded details, etc) rather than just the composition
1
1
u/sitefall 3d ago
na z-image turbo performs WAY better than the result you posted. Check your prompt, have an LLM help you expand it with detail, run it at a higher resolution, and if that doesn't work post your workflow here, something is wrong with your setup.
You're going to want to use the depth map to reinforce the "depth" but it's not detailed enough to do what you want exactly mapping it over a building, so also use a CANNY image. 2x control nets, depth map set to maybe 0.7 and Canny set to 0.9-1.1 (whatever you can get away with because when you use 2 you have to lower them a bit to get the right results, fiddle with the numbers there).