Quick Summary: This verified Google Gemini & Midjourney v6 prompt creates an authentic 90s vintage film aesthetic for Giant-scale Fashion Photography AI Photo Prompt with calibrated lighting, authentic texture physics, and tested 1-click copy-paste formula for instant AI generation without facial distortion.

- Illustration 3d
- Fashion
- Aesthetic
- Ai Image
- Trending
- Illustration 3d
- Fashion
- Aesthetic
- Ai Image
- Trending
- Viral Prompt
- Gemini Ai
- Bing Image Creator
- Copy Paste
About This Prompt
Design a Giant-scale Fashion Photography AI Photo Prompt — a high-quality, photorealistic image with balanced lighting, sharp focus, and intricate details suitable for social media sharing and digital content creation.
We recommend generating this using Google Gemini. Simply copy the prompt text below and paste it into the generator. You can easily customize specific keywords like colors, backgrounds, or subjects to make the final output uniquely yours.
What You'll Get
- 🎨 Style: Illustration 3d
- ✨ Mood: Illustration 3d
- 📱 Best For: Fashion
- 🤖 Works With: Google Gemini
- 🔥 Popularity: Copied 17 times
- 📅 Added: Sep 14, 2026
Giant-scale Fashion Photography AI Photo Prompt - Copy Paste AI Prompt
1-Tap Interactive CustomizerEpic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements.
Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting.Copy-Ready Prompts for Google Gemini, Bing & Midjourney
Different AI image generation models use different token parsers. Select your AI engine below to copy the exact formula tuned for that model's architecture:
Generate a hyper-realistic photographic image based on this description: Epic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements. Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting.. Camera & Lighting: Captured with an 85mm f/1.4 Professional Prime Lens at f/1.8 for dramatic subject pop with smooth natural depth-of-field, ISO 100 for clean noise-free shadow gradients and crisp detail, 1/200s (freeze subtle movement with optical sharpness). Lighting: Directional softbox key light at 45 degrees, subtle ambient fill, warm hairline kicker for edge definition. Color Science: Kodak Portra 400 color science, natural skin tone reproduction, balanced highlight rolloff, zero waxy plastic smoothing. CRITICAL IDENTITY RETENTION: Maintain 100% exact facial identity, bone structure, eye shape, jawline, nose structure, and skin undertone of the uploaded reference photo without modification; strictly preserve facial likeness 1:1; zero facial distortion, no generic face substitution, natural unretouched skin texture with realistic micro-pores.
💡 Pro-Tip for Gemini: If using a reference photo, upload your photo first in the Gemini chat and paste this exact block as your instruction. Gemini will seamlessly retain your facial features.
100% Face Consistency & Identity Lock Guide
The #1 complaint with AI photo generators is "mera chehra badal gaya" (facial distortion or artificial substitution). Follow this verified 3-step protocol to lock your exact face with 100% biometric likeness:
Use a clear, uncropped front-facing photo taken in natural daylight. Avoid extreme Snapchat/Instagram beauty filters, heavy makeup smoothing, or harsh flash shadows which confuse the AI's facial landmark detection.
Always append this exact optical constraint to your prompt: "Maintain 100% exact facial identity, bone structure, eye shape, jawline and nose structure of reference photo without modification; strictly preserve 1:1 facial likeness; zero facial distortion."
In Midjourney, keep --cw 100 (Character Weight 100). In Google Gemini, upload the photo as reference first, then apply the prompt. In Flux / SDXL, utilize ControlNet IP-Adapter FaceID with weight set to 0.85–0.90.
🔒 Rule 12 Guarantee: Upload a clear front-facing selfie and instruct the AI: "Retain 100% facial identity, bone structure, eye shape and skin undertone of reference photo without distortion."
Prompt Architecture & Optical Blueprint
To achieve photorealistic, non-artificial generations on Google Gemini (Imagen 3), Bing Image Creator, or Midjourney, each constituent visual token is calibrated to real-world optical physics rather than generic quality buzzwords.
Striking photorealistic subject, lifelike expression, natural skin pores, compelling eye contact, cinematic composition
Identity anchor maintaining consistent facial anatomy and posture.High-texture tailored apparel with realistic fabric draping, authentic threading, clean modern aesthetic styling
Tactile fabric micro-textures avoiding synthetic plastic sheen.Cinematic environmental setting with realistic atmospheric depth, soft specular background bokeh, natural ambient separation
Multi-plane background depth with realistic optical falloff.85mm f/1.4 Professional Prime Lens (Full-Frame 35mm Digital Sensor (50MP+ Resolution))
Simulated focal length preventing wide-angle facial distortion.Directional softbox key light at 45 degrees, subtle ambient fill, warm hairline kicker for edge definition
45-degree key light with rim kicker creating authentic cheekbone depth.Kodak Portra 400 color science, natural skin tone reproduction, balanced highlight rolloff, zero waxy plastic smoothing
Balanced highlight rolloff with natural skin micro-pore retention.Aesthetic Variations & Style Modifications
Take the core concept of Giant-scale Fashion Photography AI Photo Prompt and explore these 4 curated artistic directions, calibrated with specialized lighting, color grading, and lens characteristics:
Warm, romantic 2800K low-angle sunlight creating radiant amber rim highlights, gentle lens flare, and golden hour atmospheric haze.
Epic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements. Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting., bathed in dramatic golden hour sunset light, low sun angle at 2800K casting long cinematic shadows, warm amber rim lighting outlining the subject silhouette, subtle golden lens flare, Kodak Gold 200 film tone, authentic atmosphere, 85mm f/1.4.
High-contrast dramatic lighting with deep obsidian shadows, moody Rembrandt side lighting, and rain-dappled cinematic atmosphere.
Epic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements. Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting., dramatic cinematic noir style, moody chiaroscuro lighting with deep pitch-black shadows and stark single-source side key light, subtle rain mist in air, high tonal micro-contrast, 35mm anamorphic lens, CineStill 800T color grade, award-winning cinematography.
Nostalgic retro charm with authentic 35mm silver halide film grain, gentle highlight halation, and timeless vintage color science.
Epic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements. Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting., authentic 1980s vintage analog film photograph, shot on vintage mechanical camera with Carl Zeiss 50mm f/1.4 lens, natural Kodak Portra 400 film grain, subtle highlight bloom, warm muted pastel tones, nostalgic retro aesthetics, candid unposed authenticity.
Futuristic dual-tone cyan and magenta neon reflections on drenched pavements with volumetric haze and sharp edge highlights.
Epic ultra-wide full-body portrait, surreal photorealistic giant-scale fashion photography, cinematic urban fantasy editorial. Use my uploaded portrait/photo as the strict facial identity reference, accurately preserving my recognizable facial structure, eyes, eyebrows, nose, lips, jawline, hairstyle, hair color, skin characteristics and overall likeness. The giant character must have my face and clearly look like me. Do not alter my identity. Create a surreal yet highly photorealistic scene in which I appear as a colossal human, hundreds of feet tall, casually sitting across the rooftops of a dense Japanese city, with surrounding architecture reduced to miniature scale beneath me. Pose and body language: Place my gigantic body centrally in the composition, seated casually on the sloping tiled rooftop of a large traditional Japanese temple. My torso leans slightly backward in an extremely relaxed, confident posture. My head tilts gently backward, chin slightly raised, with eyes softly closed or nearly closed, conveying calmness, confidence and effortless presence. Both arms extend loosely outward to the sides, elbows naturally bent, hands hanging relaxed near the camera. One leg extends much farther forward and downward toward the lens, creating extreme perspective exaggeration. The sole and front of one boot become one of the closest elements to the camera and therefore appear enormous. The opposite leg bends outward and rests naturally across the architecture. The pose should communicate the impression that an impossibly gigantic version of me is casually lounging across an entire city block. Clothing: Dress me in a bold luxurious bright-red faux-fur jacket, oversized and highly textured, with a wide plush collar and thick soft fibers catching the daylight. Underneath, only a subtle glimpse of a neutral dark shirt should be visible. Wear a loose light-wash blue denim jeans, relaxed wide-leg silhouette, realistic heavy denim folds around the knees, thighs and ankles. Complete the outfit with large worn brown suede Chelsea boots, chunky soles, realistic distressed leather/suede texture, subtle dirt and natural wear. The clothing should remain physically believable at monumental scale, with realistic gravity, folds, stitching, fabric thickness and shadows. Environment: Set the scene in Tokyo, Japan, combining dense contemporary city architecture with a large traditional Japanese Buddhist temple directly beneath my gigantic body. The temple features dark gray ceramic tiled roofs, red structural columns and traditional Japanese architectural details. My giant body sits directly across the temple rooftop as if the building were furniture. Surround the temple with dense urban streets, mid-rise buildings, intersections, pedestrian crossings, sidewalks, tiny vehicles and extremely small pedestrians far below, all reinforcing my enormous scale. In the distant skyline, include Tokyo Tower on the left side and Tokyo Skytree on the right side, separated across the horizon and visually framing my gigantic body. The city must remain highly detailed and geographically believable, but the giant character is unquestionably the visual focus. Composition and camera: Vertical environmental portrait, extreme high-angle aerial perspective combined with an ultra-wide lens. The camera above street level and slightly above/beyond my extended feet, looking diagonally toward my gigantic seated body and the city behind me. My head remains around the upper-central portion of the frame, torso centered, arms expanding horizontally. torso centered, arms expanding horizontally. While my enormous legs create strong diagonal leading lines from the center toward the bottom corners. One boot enters prominently into the lower foreground, appearing dramatically larger because it is significantly closer to the camera. Use exaggerated but optically believable 14-16mm ultra-wide perspective, emphasizing spatial depth and scale without creating circular fisheye distortion. The city streets and tiny vehicles below should provide powerful visual scale references while maintaining substantial depth of field so my face, clothing, temple rooftop and surrounding city remain recognizable and detailed. Lighting and color: Bright natural daytime illumination under a soft blue sky, producing natural highlights across my face, red jacket, jeans and rooftops. The red faux-fur jacket should become the strongest color accent in the scene, contrasting dramatically against cool blue sky, muted gray architecture, pale blue denim, dark temple roofs and brown boots. Realistic atmospheric perspective should slightly soften distant buildings while foreground elements remain crisp. Natural skin tones, realistic highlight roll-off, subtle shadows beneath the body and clothing, physically believable ambient bounce light from surrounding architecture. Scale realism: The most important visual concept is extreme monumental scale. I should appear approximately 100-150 meters tall, while buildings, streets, cars and pedestrians retain normal proportions. My body must convincingly interact with the architecture: believable contact points between my body and rooftop, realistic cast shadows across the temple, proper occlusion, consistent perspective and physically coherent lighting. The result should feel like a genuine aerial photograph of an impossible event, not a miniature diorama and not an obvious Photoshop composite. Camera language: Photographed with the visual characteristics of a full-resolution medium-format camera, ultra-wide 15mm equivalent lens, approximately f/8-f/11, ISO 100, deep environmental focus, extremely detailed urban textures, realistic optical perspective and subtle photographic distortion. Surreal luxury fashion editorial meets monumental cinematic photography, sophisticated absurdism, contemporary conceptual photography, playful scale manipulation, photorealistic urban environment, sculptural natural stone pavement, moss-covered surfaces and subtle reflective water features. Hyper-accurate facial likeness to my uploaded portrait/photo reference, recognizable face, natural skin pores and microtexture, realistic hair strands, realistic clothing fabrics, individually visible faux-fur texture, detailed denim weave, weathered suede boots, realistic Japanese architecture, miniature vehicles and pedestrians for scale, physically accurate shadows, HDR, cinematic depth, crisp photographic detail, subtle film grain, premium editorial color grading, ultra-photorealistic, 8K finish. Negative prompt: different person, generic face, altered facial identity, inaccurate facial features, changed hair color, changed hair face, distorted face, duplicated person, multiple giants, extra person, extra arms, extra legs, extra fingers, malformed hands, malformed feet, warped body, broken anatomy, miniature sitting person, normal-sized person, floating body, body clipping through buildings, incorrect architectural interaction, miniature human, normalized human environment, toy city, miniature diorama, fake scale, poorly composited giant, floating architecture, distorted buildings, warped Tokyo Tower, warped Tokyo Skytree, fisheye circle, extreme barrel distortion, cartoon, anime, illustration, 3D render, video game graphics, plastic skin, waxy face, excessive beauty retouching, blurry facial details, low-resolution, oversaturated colors, artificial lighting, text, letters, words, captions, typography, title, subtitle, logo, brand name, signature, watermark, poster text, graphic overlay, UI elements. Optical Identity & Likeness Preservation: Lock subject's authentic facial identity and natural posture while seamlessly integrating with sharp vehicle reflections and cinematic ambient lighting., futuristic cyberpunk night scene, drenched wet asphalt reflecting vibrant cyan and magenta neon signage, volumetric mist illuminated by streetlights, sharp rim lighting, shot on 50mm f/1.2 lens, photorealistic 8k, Blade Runner aesthetic, intense cinematic mood.
Common AI Generation Artifacts & Forensic Fixes
If your generated image exhibits digital artifacts, unretouched plastic skin, or limb occlusions, use this diagnostic matrix to tune your generation parameters:
| Visual Artifact / Defect | Underlying Engine Cause | Forensic Fix & Prompt Modification |
|---|---|---|
| Plastic / Waxy Skin (Over-Smoothed) | Over-weighted beauty tokens ("flawless skin", "8k") triggering aggressive denoising in latent diffusion. | Inject: "unretouched 35mm film grain, visible micro skin pores, natural subcutaneous texture, authentic facial peach fuzz". Lower CFG scale from 7.0 to 4.5–5.5. |
| Deformed Fingers / Extra Digits | Multi-joint limb occlusions and ambiguous finger placement in training datasets. | Add negative tokens: (mutated hands, fused fingers, extra digits:1.4). Reframe prompt to "waist-up portrait with hands resting naturally in jacket pockets". |
| Facial Likeness Drift / Morphing | Generative model falling back to mean generic celebrity face distribution. | In Midjourney, enforce --cw 100. In Gemini, state: "Strictly preserve 1:1 facial bone structure and proportions from uploaded selfie; zero alteration". |
| Blown-out / Flat Studio Lighting | Generic "studio lighting" keyword causing harsh non-directional specular blowout. | Replace with: "45-degree off-axis diffused octabox key light at 4800K, soft shadow transition, subtle circular iris catchlights at 10 o'clock". |
Explore Related Prompt Architectures & Tools
Supercharge your workflow by building custom variations or testing complementary categories:
Explore Related Collections
Explore Illustration 3d Prompts — Browse our curated collection of Illustration 3d prompts tailored for creators. View all Illustration 3d prompts →
Explore Fashion Prompts — Discover more viral Fashion ideas for your next AI generation. View all Fashion prompts →
Explore Aesthetic Prompts — Explore our extensive library of Aesthetic templates ready to copy and paste. View all Aesthetic prompts →
Frequently asked questions
- How do I use the Giant-scale Fashion Photography AI Photo Prompt prompt?
- To use this prompt, simply click the "Copy Prompt" button above. Then, open Google Gemini, paste the copied text into the prompt box, and hit generate. You can instantly generate high-quality images without needing any reference photos.
- What is this prompt best used for?
- This specific prompt is highly optimized for creating engaging social media content and digital art. The formula is structured to produce high-contrast, visually appealing results that perform well on visual platforms.
- Can I customize the details of this prompt?
- Yes! While the core structure of the prompt is optimized for Google Gemini, you can easily customize the details. We recommend changing specific keywords like clothing colors, background locations, or the overall mood (e.g., changing "sunny" to "rainy" or "neon") to match your exact vision.











