Confident girl in retro Japanese street snap AI Image Prompt
Video prompt (Douyin trending style, confident protagonist, fast pace). Real cinematic texture, 4K HD, post-production overlay with strong soft light + halo + light gold haze + vignette + grain. Fixed camera position, soft light fade out after character action, overall oil-painting-like languid vintage. Scene setting: Japanese-style retro small shop sign "酒想见你" (Sake Miss You). Off-white textured wall (rough sand grain texture), with black artistic font "酒想见你" above, and small text below: "朋友常约酒 关系更长久" (Friends often drink, relationship lasts longer); the wall is decorated with fresh green branches (branches extending from the top of the frame, leaves fresh and lively); retro street, background is a two-story beige building with "HOUSE" sign. The upper part of the building is covered with greenery (vines hanging down, potted plants, lush green leaves), and indoor furnishings are visible through the windows (warm lighting, wooden furniture, books, etc.). The upper left corner of the frame is filled with branches and leaves of a big tree, dense foliage, sunlight casting dappled shadows through the leaves. Several retro bicycles are parked beside the building (fresh colors like mint green, cream white). There are 2-3 pedestrians walking normally on the street, serving as background to highlight the protagonist. The ground is light-colored stone bricks or asphalt road. Overall atmosphere is "fresh and artistic city street snap, full of spring/summer romance and relaxation." Passersby setting (2-3 people, as background support, not stealing the spotlight): · Pedestrian behavior: They walk, chat, or mind their own business naturally, not deliberately looking at the camera, maintaining a natural daily life state, as background to enhance the protagonist's presence. Character setting: Young woman, 22-26 years old, fair skin, delicate and three-dimensional facial features, with a natural confident smile. Long brown hair, one side braided into a delicate three-strand braid or fishtail braid from ear to ends, the rest naturally flowing, slightly curled ends, fluffy and shiny. Makeup is clear and delicate: natural eyebrows, light brown eyeshadow, defined lashes, pink-orange blush, glossy glass lip glaze. Overall temperament: gentle, romantic, relaxed. Wearing the same outfit as in the reference image. Action sequence (about 10 seconds, brisk and smooth rhythm): 0-1 sec: The protagonist squats down, hands on knees, face extremely close to the camera (almost touching the screen), delicate makeup clearly visible – clear base, light brown eyeshadow, pink blush, pink-orange glass lips, eyes confident with smile directly at the camera, as if interacting with the audience, full charisma. 1-2 sec: The protagonist quickly stands up while confidently stepping backward to a mid-back position (full body in frame, proper headroom and footroom), standing upright with straight back, poised and confident. 2-3 sec: Right hand smoothly sweeps hair back, movement crisp and clean, hair lifts lightly, eyes confidently look at the camera, showing a lazy yet powerful aura. 3-4 sec: Confident turn, body quickly turns to one side, striking the first photo pose – one hand on hip, body slightly sideways, weight on one leg, leg lines stretched, expression confident smile, frozen moment with protagonist's full charisma. 4-5 sec: Quick pose change, body turns to the other side, striking the second photo pose – one hand lightly on head or touching hair ends, the other hand naturally hanging or on waist, body forming S-curve, elegant and confident. 5-6 sec: Another quick change, slightly side-turned to camera, striking the third photo pose – one hand making peace sign ✌️, the other lightly on waist or cheek, eyes glancing at self in the mirror, expression natural and confident. 6-7 sec: Body returns to center, hands naturally fall, lightly touching collar or skirt, showing clothing details (off-white loose knit cardigan, light apricot slip dress texture, cut, color), movements smooth without dragging. 7-8 sec: Confidently toss hair, turn to face camera, cross arms or naturally place hands on hips, striking the fourth photo pose – body slightly sideways, one foot lightly tapping the ground, chin slightly raised, eyes confident and determined, strong aura. 8-9 sec: Quickly walk toward the front of the camera (take 2-3 steps forward), closing distance with the audience, stop and slightly turn sideways, one hand naturally hanging, the other lightly touching collarbone or necklace, showing confident smile. 9-10 sec: Finally strike a grand final photo pose – body slightly sideways, one hand on hip, the other naturally hanging, legs stretched, chin slightly raised, eyes confidently looking at the camera, freeze, soft light fade out. Light and color tone: Daytime natural light, soft sunlight (golden hour before 10am or after 3pm), big tree leaves cast dappled shadows on the upper left of the frame, building greenery intertwines with sunlight. Light falls from the upper side, gently wrapping the character's face and clothing, skin appears fair, clothing texture stands out. Filter: Strong soft light + halo + light gold haze + vignette + grain, soft warm tone (slightly creamy/filmic), low saturation high softness, clothing colors remain true. Final: Character skin looks fair, clothing texture, cut, and color clear, light and shadow soft and beautiful, every frame can serve as product image, strong purchase desire.
Content: Prompt structure: overall style + scene setting + character setting + action sequence (per second) + lighting & color tone + filter effect. Subject: young woman performing a 10-sec continuous posing action in front of a retro street. Scene includes Japanese shop, greenery, bicycles. Style: cinematic, soft light, vintage. Composition: fixed camera. Lighting: natural light with post-production soft light & haze. Color: warm, low saturation. Texture: rough wall vs. clear makeup. Lens: fixed.
Pros: 1. Action sequence detailed per second, clear rhythm, easy for AI to understand timing; 2. Specific scene and lighting descriptions help generate consistent visuals.
Cons: 1. Prompt too long, may exceed model context limits; 2. Relies on reference image to fully reproduce outfit.
Reference image: Clear dependency on reference image (prompt mentions 'wearing the same outfit as in the reference image')
Source materials are collected from publicly accessible web pages. Contact us if a rights issue needs review. Copyright & Privacy