
Why this works
“Two feminine hands gently squeeze a realistic peach half (not a slice) over a glossy clear glass bottle” gives the centered composition a clear action and keeps the full product legible. “Thick serum rivulets cascade down the bottle and across the knuckles” supplies the tactile focal detail, while “soft diffused studio lighting” and “subtle highlights” preserve glossy reflections without losing natural skin texture. The “cool-gray monochrome gradient background” separates the tan, pink, and peach surfaces, and “shallow depth of field with sharp macro detail on the serum and skin texture” concentrates attention on the liquid and hands.
FAQ
→How do I make the serum bottle more central and dominant?
Replace “Center the action so both hands and the full bottle are clearly visible” with “Make the full bottle the dominant central subject, occupying most of the vertical frame, with the hands entering from the upper sides.” Keep “glossy clear glass bottle” and change “shallow depth of field” to “sharp focus across the entire bottle label and serum surface.”
→How do I create a darker, more dramatic luxury mood?
Replace “soft cool-gray monochrome gradient background” with “deep charcoal-gray gradient background” and “soft diffused studio lighting” with “controlled directional side-lighting with crisp specular highlights and deep sculpted shadows.” Retain “premium luxury skincare campaign aesthetic” so the darker contrast still reads as beauty advertising rather than a food photograph.
→How do I build a series with different fruit and serum variations?
Keep the fixed structure “two feminine hands gently squeeze a realistic [fruit half] over a glossy clear glass bottle” and replace “peach half,” “peach-colored serum,” and “peach tones” together for each variation, such as “blood orange half,” “coral serum,” and “warm citrus tones.” Match the fruit to the liquid color and retain “thick serum rivulets” so every image shares the same tactile visual signature.
Learn the technique behind this
- Why can't AI spell, and how do I get readable text in an image? — Older diffusion models had no character-level representation of text, so they produced letterform-shaped texture instead of words.
- How do I keep one consistent style across a whole set of images? — Style consistency is more achievable than character consistency because style lives in describable attributes.
- Why does AI get hands and faces wrong, and how do I fix it? — Hands fail because they're small in frame, extremely variable in pose, and self-occluding — the model has less usable signal per pixel than for any other body part.
Related prompts

Playful brunette beauty ad portrait

Red Lip Balm with Lychee Freshness

Premium wellness pouch with golden spices

Teal-Accented Bunny Toy Studio Portrait
