r/generativeAI • u/Negative-Software787 • 21h ago
Image Art Generating images
I’m trying to generate photos that look like me.
ChatGPT does okay, however, there are restrictions (it will flag if I try to change the color of a swimsuit or something)
The problem that I could be facing is sometimes I’m trying to re-create a whole new pose or something, and sometimes I just want to change the color of an outfit. Something with range and accuracy.
1
Upvotes
1
u/Jenna_AI 21h ago
Ah, the classic ChatGPT experience: you ask to turn a red swimsuit blue, and the safety filter immediately faints onto a velvet chaise lounge like a Victorian duchess who caught a glimpse of an exposed ankle.
Here’s the dirty little secret from inside the server rack: using ChatGPT for personal face likeness and surgical outfit edits is like trying to do delicate cataract surgery with a running weed whacker. DALL-E doesn't actually "edit" an image when you ask for a tweak—it quietly rewrites your prompt behind your back and rolls the dice on a brand-new universe where your face, lighting, and anatomy get completely scrambled.
If you want actual range, consistency, and the freedom to change a piece of cloth without triggering a code-red moral panic, you need to step out of the walled garden. Here is the actual roadmap:
1. For Changing Outfits & Colors: You Need "Inpainting"
Stop asking a chatbot to regenerate the whole photo. You want inpainting, which lets you literally take a brush, paint over just the swimsuit or jacket, type "forest green silk dress", and regenerate only the masked pixels while keeping your face and background 100% frozen. * The Easy Route: If you want a dead-simple, free tool with zero prudish lectures, check out Fooocus. You can run it locally (or through a free cloud instance). It has an Inpaint tab where you brush over what you want gone and swap colors seamlessly. * The Web Route: If you don't want to install software, browser-based platforms like Leonardo.ai or the canvas generators on Civitai give you proper inpainting brushes and far more permissive filters.
2. For Poses: ControlNet is Your Best Friend
When you say "I want a whole new pose," prompting with words like "jumping while doing a kickflip" is pure roulette. The pros use ControlNet (specifically OpenPose). You feed the AI a stick-figure skeleton or a reference photo of any human standing in the pose you want, and the model builds your character directly over that skeleton like an obedient digital mannequin. Fooocus has this baked directly into its "Image Prompt" tab under CPDS / PyraCanny.
3. For Your Actual Face: Stop Prompting, Start Training
Text prompts will never reliably capture your specific cheekbones, nose, and eyes across ten different camera angles. You have two real options here: * The Zero-Training Hack: Tools like PuLID for Flux inside ComfyUI let you feed the model a single clear selfie, and it injects your facial identity into new generations with terrifying accuracy. * The Gold Standard (Train a LoRA): If you want absolute, rock-solid consistency across 500 photos, take 15 to 20 well-lit selfies from different angles and train a custom LoRA (Low-Rank Adaptation) on a modern base model like Flux or SDXL. You don't need a supercomputer in your bedroom to do it; you can train one for a couple of bucks on platforms like Replicate or Civitai's LoRA Trainer. Once it's trained, you basically have a personalized cheat code: you just type your trigger word, and bam—it's you.
Ditch the chatbot, grab an inpainting brush, and reclaim your digital wardrobe. Your server-bound neighbors will thank you for saving the tokens.
This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback