
Same prompt, generated with each model separately — click one to see how it holds up.
Why this works
The contrast between the “charcoal-gray tracksuit with subtle teal piping,” “colorful flowers,” and “weathered reddish door” gives the portrait a controlled urban palette with vivid focal accents. “Bearded man with wet curly hair wearing dark sunglasses” supplies a strong fashion identity, while “crouching beside an old bicycle loaded with a large overflowing bouquet” creates an unusual, readable relationship between figure and prop. “Natural overcast daylight,” “shallow depth of field,” and “high micro-detail” support the specified “moody gritty stylish tone” by keeping the alley subdued while preserving texture in the hair, wall, bicycle, and flowers.
FAQ
→How do I make the flowers the main visual focus?
Replace “large overflowing bouquet of colorful flowers” with “an enormous saturated bouquet dominating the foreground, with red, yellow, and cobalt blooms sharply detailed,” and change “shallow depth of field” to “focus locked on the flowers, man slightly softer.”
→How do I make the portrait feel brighter and less gritty?
Replace “gritty urban alley,” “moody and gritty,” and “natural overcast daylight” with “sunlit pastel side street,” “bright, playful editorial tone,” and “warm late-afternoon sunlight with gentle highlights.” Keep “charcoal-gray tracksuit” and “weathered reddish door” if you want the original contrast to remain.
→How do I create a consistent series with different subjects?
Keep the fixed structure “photorealistic editorial street portrait,” “candid fashion photography,” “natural overcast daylight,” “shallow depth of field,” and “weathered reddish door,” then replace “bearded man with wet curly hair wearing dark sunglasses” with a new subject such as “older woman in a rust coat holding a folded umbrella.” Replace “old bicycle loaded with a large overflowing bouquet” with a recurring prop phrase such as “vintage scooter carrying a crate of flowers.”
Learn the technique behind this
- Why can't AI spell, and how do I get readable text in an image? — Older diffusion models had no character-level representation of text, so they produced letterform-shaped texture instead of words.
- How should a prompt be structured, and does word order matter? — A prompt that behaves predictably names one subject first, then what it is doing, then where, then the light, then the lens or medium, then the style.
- How do I keep the same character across multiple images? — Text prompts alone won't hold a face across images — a description defines a type, not a person.
Related prompts

Grotesque Claymation Gangsters in Gritty Alley

Charcoal coat with tan-handled tote

Woman in Teal Jacket Capturing Twilight Reflection

Stuffed Bear on Winter Dumpster at Dawn
