
Why this works
The phrase “photorealistic full-body fashion portrait” preserves the complete outfit and gives the image a clear vertical structure, while “standing on worn outdoor concrete stairs in a narrow urban alley” adds layered gray-brown geometry around the subject. “Golden-hour street lighting” supplies the warm nostalgic mood, and “slightly nostalgic film-like look” shapes the muted brown, gray, teal, and blue palette rather than letting the solid emerald blouse dominate unnaturally. “Hands in pockets,” “relaxed pose,” and “candid street photography” keep the fashion styling informal, while “shallow depth of field” separates the woman from the alley without losing the recognizable stair setting.
FAQ
→How do I change this prompt to make the woman feel more confident and editorial?
Replace “relaxed pose with hands in pockets” with “confident upright stance, one hand on hip, direct gaze into the camera,” and replace “candid street photography” with “polished editorial fashion photography.” Keep “full-body fashion portrait” to retain the complete outfit and stairs.
→How do I make the emerald blouse the main visual focus?
Change “a solid emerald green wrap-style blouse” to “a vivid jewel-toned emerald green wrap blouse, high color contrast and sharply defined fabric,” and add “subject sharply focused, background heavily softened.” You can also replace “shallow depth of field” with “very shallow depth of field” so the gray-brown alley recedes more strongly.
→How do I build a matching series with different urban locations?
Keep “young woman,” “shoulder-length auburn hair,” the blouse, jeans, boots, “full-body fashion portrait,” and “slightly nostalgic film-like look” unchanged, then replace “narrow urban alley” and “worn outdoor concrete stairs” with locations such as “sunlit tiled subway entrance,” “weathered brick courtyard,” or “rain-darkened rooftop access.” Retain “golden-hour street lighting” across the series for consistent warmth and color treatment.
Learn the technique behind this
- Why does AI get hands and faces wrong, and how do I fix it? — Hands fail because they're small in frame, extremely variable in pose, and self-occluding — the model has less usable signal per pixel than for any other body part.
- Why can't AI spell, and how do I get readable text in an image? — Older diffusion models had no character-level representation of text, so they produced letterform-shaped texture instead of words.
- How do I keep the same character across multiple images? — Text prompts alone won't hold a face across images — a description defines a type, not a person.
Related prompts

Teal fashion editorial in monster-filled lift

Cobalt Boots in Sleek Studio Edit

Silhouette Influencer by Sunlit Window

Sunlit editorial on striped beach towel
