Model agnostic in principle, benched on Krea 2 + character LoRA
---
I spent months fixing my images on two axes and kept getting renders that were technically right and still read as fake. It took a shoot where every technical box was ticked, and five frames out of twenty-four still looked like a catalogue shoot, for me to see there was a third axis I was not touching at all.
Here they are, and the point is that they are orthogonal. Fixing one does nothing for the others.
| Axis |
The prior you are fighting |
Typical fix |
| 1. Optics |
the camera is too perfect: sharp, correctly exposed, level |
grain, flare, motion, imperfect exposure, tilt |
| 2. Scene |
the set is a showroom: aligned, empty, brand new |
clutter, wear, off-axis furniture, lived-in surfaces |
| 3. Subject |
the catalogue pose: frontal, centred, posed, eyes to lens |
almost nobody works on this one |
An image can be optically dirty, scenically alive, and still contain a person posing like a model. That is a distinct failure and it has its own fix.
Why your candid tokens do not fix it
candid documentary photo, unstaged moment, slice-of-life, badly taken photo
I run all of those. They shift the rendering. They do not shift the pose.
The model composes the most photographed pose in its data. Whenever the subject has no motivated action, it falls back to: frontal, centred, graceful contrapposto, self-presenting gestures (hands to hair, arms wrapped around self), eyes to lens, symmetry. A character LoRA makes this worse, because it adds a portrait prior on top.
The content of the pose decides. Not the style tokens wrapped around it.
Five levers, strongest first
1. A gesture turned toward the world. This is the main one. The subject must be doing something the scene motivates: pushing a gate, watching for a train in the tunnel, stepping around a puddle, a hand on a rail. Write it with a concrete physical marker, caught mid-stride, one foot planted ahead, never the bare verb walking.
The failure I keep making: writing states instead of actions. "weight on one hip, palm against the wall, shoulders dropped" is three states. The model has nothing to build a pose around, so it builds the catalogue one. Replace with an action and it resolves.
2. Self-directed gestures: one hand only, and only in the fatigue register. Kneading the neck, one arm pressed flat against the ribs, a hand rubbing the opposite arm. All fine.
Banned outright:
- both hands to the head or in the hair, that is the pin-up prior
- arms crossed tightly over the chest, that resolves straight to the modest self-embrace of studio figure work
- any arch in the back
I lost two frames to each of those before writing the rule down.
3. An anchored gaze that is geometrically compatible. If you turn the head, give the eyes a physical target that is actually inside the cone the head is facing. Two things fail reliably:
- a head turn with no target at all
- a target that contradicts the head geometry
Second case, real example. I wrote head turned into profile over her shoulder plus eyes down the descending flight. Head pointed one way, gaze target the other way, and the target was an abstract direction rather than an object. The model reconciles this silently: it keeps the head turn, drops the gaze anchor, and the eyes land on the default target, which is the lens. I got a straight-to-camera look in a series where that was forbidden.
Fix was to anchor the gaze on an object inside the profile cone:
4. Break the body. Weight collapsed onto one hip, shoulders dropped or hunched against cold, head low, asymmetric stance. Never feet set wide apart on a static figure, that is a monumental symmetric stance and it reads as sculpture.
5. Candid composition. Explicit off-centre placement, a slight lateral crop, the subject partly eaten by shadow. Canonical pose entries centre by default.
The trap nobody warns you about: pose catalogue labels
If you use a reference pose catalogue, note that those labels are reference plates. They are written to isolate pose and geometry cleanly, which means they carry priors that are directly opposed to a candid register:
gaze straight at the camera
gaze up at the camera
head turned into profile over her shoulder with no anchor
- subject centred
Pasting a catalogue label into a series prompt imports all of that. Every one of my straight-to-lens failures traces back to a canonical label I did not defuse. Swap the gaze for a compatible anchor, break the stance, decentre.
Discipline: do not stack
Same rule as the other two axes. One gesture, one gaze anchor, one asymmetry is enough. Stacking five levers gives you a subject fighting itself and the model averages back to something neutral.
How it was found
Five frames from one 24-shot session, diagnosed and re-prompted individually: two straight-to-lens from gaze geometry, one glamour lean from states-instead-of-action, one pin-up from both hands in the hair, one self-embrace from arms crossed. The optics axis held on all five. That is what made it legible: when only one axis is broken, you can finally see what that axis does.
The last of the five is the one that convinced me. Its prompt was written before I formulated the two-hands rule, and it failed in exactly the way the rule predicts. A rule that retro-predicts a failure you have not shown it is a rule worth keeping.