
Why this works
The first-person perspective and “sharp focus on the holographic phone,” supported by “shallow depth of field,” make the hand-held interface the visual anchor while the wide street recedes around it. “Holograms shifting from cyan-blue to violet accents” and “high-contrast violet-and-cyan neon lighting” create a controlled purple-and-cyan harmony against the dominant black cityscape, while “light rain and wet asphalt with crisp reflections” carries those colors across the ground. “Futuristic skyscrapers with dense neon signage,” the “humanoid robot,” and the “drone-like aerial vehicle hovering overhead” establish scale and a dramatic high-tech mood without competing with the phone’s sharp focus.
FAQ
→How do I make the humanoid robot more central and prominent?
Replace “a humanoid robot walking along the curb” with “a humanoid robot standing in the center of the street, full body clearly visible, sharp focus, illuminated by violet-and-cyan neon.” Replace “sharp focus on the holographic phone” with “sharp focus shared between the holographic phone and the humanoid robot” to reduce the phone’s dominance.
→How do I turn this into a warmer, less ominous cyberpunk scene?
Replace “high-contrast violet-and-cyan neon lighting” with “warm amber, magenta, and soft cyan neon lighting,” and change “dramatic sci-fi city ambience” to “welcoming late-night futuristic city ambience.” Replace “black” or the dark-heavy tonal direction with “rich midnight blue shadows with warm reflected light” to preserve night depth without making the street feel severe.
→How do I create a series of variations while keeping the same visual identity?
Keep “first-person perspective,” “holograms shifting from cyan-blue to violet accents,” “wet asphalt with crisp reflections,” “ultra-detailed photorealistic rendering,” and “shallow depth of field” unchanged. Swap only the scene anchors, such as replacing “a humanoid robot walking along the curb” with “a masked courier beside a noodle stall,” then “a sleek drone-like aerial vehicle hovering overhead” with “a monorail passing between the skyscrapers,” while retaining “sharp focus on the holographic phone” across the set.
Learn the technique behind this
- Why can't AI spell, and how do I get readable text in an image? — Older diffusion models had no character-level representation of text, so they produced letterform-shaped texture instead of words.
- Why does AI get hands and faces wrong, and how do I fix it? — Hands fail because they're small in frame, extremely variable in pose, and self-occluding — the model has less usable signal per pixel than for any other body part.
- How do I keep the same character across multiple images? — Text prompts alone won't hold a face across images — a description defines a type, not a person.
Related prompts

Grotesque Claymation Gangsters in Gritty Alley

Charcoal coat with tan-handled tote

Woman in Teal Jacket Capturing Twilight Reflection

Stuffed Bear on Winter Dumpster at Dawn
