LogoAhaPrompt
  • Prompts
  • Featured
  • Models
  • Submit
  • Blog
AhaPromptAhaPrompt

AI Prompts for Images, Video & Audio — Every Model, One Library

Built withLogoNEXTY.DEV
Explore
  • All Prompts
  • Featured Prompts
  • Image Prompts
  • Video Prompts
  • Audio Prompts
  • AI Models
  • Submit a Prompt
  • Blog
Languages
  • English
  • 中文
  • 日本語
  • 한국어

Copyright © 2026 AhaPrompt All rights reserved.

Privacy PolicyTerms of Service
HomeAI Prompt LibraryAI Image PromptsAnime World Miniature Diorama Sculpture
Anime World Miniature Diorama Sculpture — AI image prompt — GPT Image 2 — Surreal, Cinematic, Fantasy - 1
Anime World Miniature Diorama Sculpture — AI image prompt — GPT Image 2 — Surreal, Cinematic, Fantasy - 2
Anime World Miniature Diorama Sculpture — AI image prompt — GPT Image 2 — Surreal, Cinematic, Fantasy - 3
Anime World Miniature Diorama Sculpture — AI image prompt — GPT Image 2 — Surreal, Cinematic, Fantasy - 4

Anime World Miniature Diorama Sculpture

OpenAIGPT Image 2OpenAIimage

Ultra-detailed photorealistic isometric miniature diorama of the world from [ANIME], presented as a premium collectible-scale landscape sculpture floating above a clean warm off-white studio background with a subtle natural shadow underneath. (For best accuracy, upload a clear reference image of the official map/world layout.) WORLD SHAPE & GEOGRAPHY If [ANIME] features a fictional world, continent, island or region, recreate its canonical map silhouette exactly. The entire landmass should appear physically carved from the earth and lifted upward as one continuous miniature terrain slab, with naturally broken rock sides, exposed soil layers, cliffs and geological textures. If [ANIME] takes place in a real or semi-real location, use the most important canonical city, region or country as the physical base shape, maintaining its recognizable geographic outline. The outer boundary must feel like an actual piece of land cut directly from the map — never like a flat printed map or floating board. CANONICAL WORLD RECONSTRUCTION Rebuild the world of [ANIME] across the entire miniature surface with maximum lore accuracy. Every major region, settlement, city, landmark, mountain, river, forest, road, battlefield, castle, shrine, futuristic structure or other recognizable location should appear in its correct relative position. Preserve the distinctive visual language of the anime: - canonical architecture - recognizable environments - authentic terrain - faction territories - signature landmarks - characteristic colors - world-specific technology or fantasy elements - environmental atmosphere Avoid generic Western architecture or realistic substitutions. The result must immediately feel like [ANIME] transformed into a physical miniature world. Add tiny but readable location markers for important places, designed like elegant museum-style cartographic labels rather than large text. MICRO-SCALE STORY DETAILS Fill the landscape with carefully placed miniature storytelling elements: tiny characters wearing recognizable show-accurate outfits, world-specific creatures, vehicles, ships, trains, flags, faction symbols, market areas, farms, bridges, pathways, trees, vegetation and other details that make the world feel alive. Characters should remain small environmental elements rather than dominating the composition. Include subtle activity throughout the terrain so the diorama feels like a living snapshot of the anime universe. MATERIAL & PHOTOGRAPHY Hyper-realistic premium miniature photography combined with high-end cinematic environment design. Extremely crisp micro-details, realistic rock formations, miniature vegetation, tiny architectural textures, believable scale relationships and physically accurate materials. Everything should remain sharply focused with no artificial background blur. Use soft warm studio illumination from above with gentle contact shadows and subtle ambient occlusion between buildings, terrain and rock layers. The miniature should have a sophisticated museum-quality collectible appearance while still preserving the anime's original visual identity. COLOR & ATMOSPHERE Use the authentic color language of [ANIME] as the primary visual reference. Translate the anime's signature palette into realistic miniature materials without losing its recognizable mood, atmosphere or world identity. Rich but controlled colors, cinematic tonal depth, premium print-like contrast and realistic surface textures. COMPOSITION Center the complete shaped diorama in the lower two-thirds of the 4:5 vertical frame. Leave generous clean negative space around the entire landmass. The diorama must never touch the frame edges. The upper third remains almost completely empty and off-white, creating a sophisticated editorial poster composition. At the very top, place small widely tracked uppercase typography showing the world / region name. Directly underneath, place: [ANIME] in a bold condensed sans-serif typeface, medium charcoal grey, large and dominant. Below it, add a smaller widely tracked uppercase line containing the studio name and year. Typography should feel like a luxury museum exhibition poster — restrained, elegant and secondary to the miniature world. FINAL ART DIRECTION Premium editorial travel-poster aesthetic + cinematic miniature photography + collectible world-building sculpture + museum-quality cartography. The final image should look as though the entire world of [ANIME] was physically sculpted from the earth and photographed inside a professional studio. No generic fantasy map, no flat illustration, no random landmarks, no duplicated buildings, no distorted geography, no excessive text, no modern Western substitutions. 4:5 vertical aspect ratio, ultra-high detail, photorealistic materials, cinematic lighting, premium print quality.

Details
Type
AI Image Prompts
Aspect Ratio
561:701
Updated
Aug 25, 2026
Category
SurrealCinematicFantasy
Source

How to use this prompt

  1. 1Copy the full prompt with the copy button above.
  2. 2Open GPT Image 2 by OpenAI, or any image generator with a similar style range.
  3. 3Paste the prompt, then swap the subject and style keywords to match your idea.
  4. 4Match the example's settings (561:701 · Surreal · Cinematic · Fantasy), then iterate — small wording changes shift the output a lot.

Frequently asked questions

Which AI model generated this image?

This image was generated with GPT Image 2 by OpenAI. The prompt on this page is the exact text used to create it.

What settings does this prompt use?

The example was generated with GPT Image 2 using: 561:701 · Surreal · Cinematic · Fantasy. Reusing these settings gives the closest match.

How do I get similar results?

Copy the prompt as-is first, then replace only the subject while keeping the style, lighting, and composition keywords. Generate a few variations — outputs differ between runs.

Can I modify this prompt?

Yes. Prompts work best as starting points — adjust descriptors, aspect ratio, or style tags to make the output your own.

More GPT Image 2 prompts

Create an ultra-realistic cinematic 4:5 vertical travel photograph of a sleek, minimalist travel car — AI image prompt — GPT Image 2
OpenAIGPT Image 2
Create an ultra-realistic cinematic 4:5 vertical travel photograph of a sleek, minimalist travel card held naturally in one hand against an expansive sky. Use [COUNTRY NAME] as the creative focus. Transform the card into a living window into the destination, with a breathtaking miniature aerial world seamlessly emerging from its surface. Feature the country's most iconic landmark, surrounded by authentic landscapes, architecture, atmosphere, and subtle details unique to the destination. Make the transition between the physical card and miniature world seamless, magical, and physically believable, as if the entire destination exists inside the card. Integrate [COUNTRY NAME] in elegant, bold uppercase typography as part of the card design. Use cinematic natural light, realistic textures, atmospheric depth, subtle reflections, dramatic perspective, soft lens falloff, and premium editorial travel-photography aesthetics. Ultra-photorealistic, 8K detail, cinematic color grading, realistic skin and materials, physically accurate lighting, luxurious, emotional, aspirational, universally beautiful. No cartoon, no illustration, no artificial CGI appearance.
Details
Create a charming miniature 3D paper-craft travel illustration of [DESTINATION], designed as a delic — AI image prompt — GPT Image 2
OpenAIGPT Image 2
Create a charming miniature 3D paper-craft travel illustration of [DESTINATION], designed as a delicate handmade travel postcard. Use only the destination name to identify the location, then feature its most recognizable local landmarks, architecture, and cultural elements as tiny handcrafted models arranged naturally on a narrow floating landscape strip. Include charming miniature details that reinforce the destination, such as authentic local transportation, trees, street lamps, pathways, water, birds, clouds, and a small airplane with a subtle dotted flight path. Keep every landmark recognizable and geographically appropriate to the destination. Use a soft white textured paper background with gentle natural lighting, pastel colors, soft shadows, and a clean airy composition. Make the miniature elements appear handcrafted from layered paper, clay, wood, cardstock, and other delicate craft materials, with tactile textures, tiny details, subtle imperfections, and realistic miniature depth. Keep the scene centered and elegant with generous negative space around the floating landscape strip. Add a small refined handwritten-style title at the bottom reading “[DESTINATION]”. Style: miniature diorama, handcrafted paper art, whimsical 3D illustration, travel postcard, soft pastel palette, tactile craft textures, macro photography feel, delicate details, cozy artistic aesthetic, realistic miniature depth, high resolution, vertical 3:4 composition.
Details
9:16 vertical format, real-life photography snapshot. An adult woman sits on an old wooden plank pat — AI image prompt — GPT Image 2
OpenAIGPT Image 2
9:16 vertical format, real-life photography snapshot. An adult woman sits on an old wooden plank path by a lotus pond in midsummer, her body slightly turned toward the camera as she leans down and reaches out to touch the water surface. Her long natural black hair is gently tousled by the summer breeze, her expression natural, as if the photographer captured a fleeting moment from very close range. She is wearing a complete, lightweight white summer outfit with white thigh-high stockings — the clothing is real and fully intact. The camera is positioned close to the subject at an extremely low angle, using a 24mm ultra-wide lens for close-range shooting, producing noticeable yet natural near-distance perspective. Giant lotus leaves, petals, and thick lotus stems intrude directly into the frame from in front of the lens, occupying a large portion of the image, extremely close to the camera and heavily out of focus, creating strong foreground obstruction. A huge pink lotus flower sits slightly above center frame, its petals just covering the key area of the subject's upper body. Lotus stems run vertically through the center of the frame, and foreground leaves drape down from the top — multiple natural elements together form a visual framing device and partial information occlusion. The woman's face emerges from between the lotus flowers and leaves, with only her clear eyes, face, and part of her hair visible. The rest of her body is naturally cut off by foreground foliage and the environment, creating a visual illusion of "seemingly missing, actually concealed." Her legs are positioned on the right side of the frame, with white thigh-high stockings clearly visible, forming a strong scale contrast with the foreground lotus flowers and leaves. Avoid deliberate posing — the movement is natural, and the body posture matches a real seated position. The background is a real summer rural lotus pond: vast green lotus leaves, blooming pink lotus flowers, shallow water reflections, an old wooden plank path, and low old rural houses and trees in the distance. The sunlight is strong but natural — harsh midday summer light, fine dappled shadows cast by lotus leaves, shimmering highlights on the water surface. The subject sits at the boundary between light and shadow, partially backlit, with genuine lens flare. Shallow depth of field: foreground lotus flowers and leaves are extremely blurred, the mid-ground subject is relatively sharp, and the background is softly and naturally blurred. Authentic snapshot quality from a phone or portable camera, natural skin texture, real hair detail, slight digital noise, delicate highlights, mild overexposure, real optical distortion, atmospheric depth, an unposed candid moment, strong spatial depth, overwhelming foreground pressure, modern East Asian summer photography aesthetics — clean and natural, high-end but not commercialized. Bottom-left corner signature "● DeepBlue," where ● is solid #0B3D91 deep blue, and DeepBlue is in white natural handwriting.
Details
Create a premium editorial art poster for every uploaded photograph, treating each image as its own  — AI image prompt — GPT Image 2
OpenAIGPT Image 2
Create a premium editorial art poster for every uploaded photograph, treating each image as its own independent composition and never merging multiple photos together. Use a strict 3:4 vertical format with the canvas split into two perfectly equal horizontal halves: the upper half should remain a faithful, photorealistic presentation of the original image, preserving the subject’s exact identity, facial features, proportions, pose, clothing, objects, composition, lighting, shadows, mood, and natural colors, enhanced only with sophisticated editorial color grading and seamless environmental extension where necessary; the lower half should transform the visual story into an entirely different artistic interpretation—a tiny, carefully composed handmade mixed-media artwork centered within expansive warm ivory negative space, occupying no more than 10–20% of the lower section, using expressive ink sketching, layered gouache-like color fields, subtle collage textures, torn-paper edges, imperfect brushwork, visible fibers, soft pigment variations, and charming human imperfections while retaining the most recognizable silhouette, gesture, objects, and emotional narrative from the original photo. Extract up to four dominant harmonious colors from each photograph and reinterpret them in a muted, sophisticated palette. Add only occasional understated editorial typography when it genuinely enhances the composition, such as a poetic title, place, date, or single word. The overall result should feel like a collectible contemporary art publication cover—minimal, poetic, tactile, elegant, emotionally quiet, visually distinctive, and unmistakably connected to its original photograph.
Details
A photorealistic cinematic countryside portrait of a cute young adult woman sitting in a lush green  — AI image prompt — GPT Image 2
OpenAIGPT Image 2
A photorealistic cinematic countryside portrait of a cute young adult woman sitting in a lush green meadow at golden hour, with the same overall composition and atmosphere as the reference. She has natural freckles, soft youthful facial features, expressive green eyes, and long wavy auburn-red hair blowing gently in the wind. She wears a colorful vintage floral dress with a flowing skirt and delicate puff sleeves. She sits naturally in the tall grass while gently holding a fluffy white rabbit beside her. Keep the same meadow, rolling green hillside, blue sky, soft clouds, warm sunset lighting, camera perspective, natural shadows, rabbit placement, and peaceful mood. Preserve the nostalgic analog-film aesthetic with subtle film grain, realistic skin texture, detailed individual hairs, natural fabric texture, authentic grass, soft atmospheric depth, and warm golden highlights. Style: ultra-photorealistic, cinematic photography, natural beauty, authentic outdoor lighting, 35mm film look, shallow depth of field, subtle grain, realistic color grading, highly detailed, editorial countryside portrait, no artificial-looking skin, no CGI appearance. Negative prompt: child, teenager, cartoon, anime, plastic skin, excessive makeup, distorted face, extra fingers, malformed hands, duplicate rabbit, distorted anatomy, oversaturated colors, blurry face, artificial lighting, CGI, text, watermark.
Details
GPT Image 2 on ChatGPT 

Prompt:

Ultra-cinematic surreal clone photography, aerial high-angle view  — AI image prompt — GPT Image 2
OpenAIGPT Image 2
GPT Image 2 on ChatGPT Prompt: Ultra-cinematic surreal clone photography, aerial high-angle view of a young Indian man standing motionless at the center of a vast urban plaza while dozens of identical copies of him move around in different directions. The clones all share the same appearance as each other, creating a striking visual repetition effect, but are not based on any reference image or real person. Handsome Indian male with a sharp jawline, neatly styled textured black hair, trimmed beard, expressive dark eyes, natural skin texture, and realistic facial proportions. Every clone has the exact same face, outfit, hairstyle, and body proportions. The central character remains perfectly sharp and isolated from the crowd, standing confidently with hands in pockets, wearing an oversized cream hoodie, black cargo pants, white sneakers, and dark sunglasses. Surrounding clones wear the identical outfit and appear slightly motion-blurred as they walk, run, and cross paths around him. Strong contrast between stillness and motion. Expansive geometric stone pavement, minimal modern architecture, subtle city elements, warm early-morning sunlight, soft atmospheric haze, realistic shadows, cinematic depth, and premium editorial fashion photography aesthetics. Tilt-shift effect creates a miniature-world illusion while maintaining realistic scale. Mood is introspective, powerful, and surreal, one person surrounded by endless versions of himself. Hyper-realistic skin detail, authentic anatomy, natural proportions, realistic clothing folds, cinematic color grading, shallow depth of field, HDR lighting, subtle film grain, luxury fashion campaign quality, contemporary surreal visual storytelling. No face distortion, no extra limbs, no duplicated body errors, no asymmetrical eyes, no cartoon style, no artificial AI-looking faces. Shot on Sony A7R IV, 85mm lens, f/1.8, aerial perspective, ultra-detailed, 8K, vertical 9:16 composition.
Details
Daily snapshot style, not carefully composed or lit, beside a bench on a school corridor, striped T- — AI image prompt — GPT Image 2
OpenAIGPT Image 2
Daily snapshot style, not carefully composed or lit, beside a bench on a school corridor, striped T-shirt with a suspender skirt, a small bear charm on the backpack, a can of Coca-Cola in hand; shot from behind a road sign, with the sign edge and leaves forming a double-layer foreground frame, slightly out of focus; overcast natural light, low contrast, low saturation, fabric texture and small accessories clearly visible, youthful and cute." "60mm, f/2.8, 1/400s, ISO 200, WB 5400K; local sharpening only on accessories and around the eyes, natural fair skin, visible clothing fabric texture, Japanese-style low saturation. Slight motion blur, overall ordinary and everyday, sweet and cute style.
Details
Ultra-realistic cinematic vintage golden-hour portrait of the same woman sitting indoors beside an o — AI image prompt — GPT Image 2
OpenAIGPT Image 2
Ultra-realistic cinematic vintage golden-hour portrait of the same woman sitting indoors beside an old, weathered wooden door. Her dark brown hair is loosely tied into a messy high bun, with a few natural strands falling softly around her face and neck. She is shown in a graceful three-quarter side-profile view, looking gently toward the left with a calm, thoughtful expression. She wears a simple oversized ribbed brown/beige knit sweater with highly realistic woven fabric texture. Warm late-afternoon sunlight enters from the side, illuminating her face, hair, and sweater in rich golden tones while casting a clearly defined natural silhouette of her head and messy bun onto the old wooden door behind her. The environment features a rustic aged interior with cracked and peeling plaster walls, distressed wooden panels, subtle imperfections, and a small vintage framed picture on the wall. Moody nostalgic atmosphere, authentic natural skin texture, realistic pores, subtle film grain, warm earthy color palette, deep shadows, cinematic contrast, volumetric golden sunlight, shallow depth of field, soft natural bokeh, editorial portrait photography, 50mm lens, f/1.8, HDR, highly detailed, photorealistic, cinematic composition, 8K. Use the uploaded reference image as the strict identity and face reference. Preserve the woman's facial identity with 100% face accuracy. Her face must remain exactly recognizable as in the reference, including the same facial structure, proportions, eyes, eyebrows, nose, lips, cheeks, jawline, skin tone, natural skin texture, and distinctive facial details. Do not modify, beautify, reshape, age, de-age, or reinterpret her face. https://t.co/HcGTzKbwBQ face replacement, no facial redesign, no beauty filter, no skin smoothing, no altered facial proportions, no changed eyes, nose, lips, eyebrows, jawline, or cheek structure. Negative prompt: cartoon, anime, CGI, 3D render, plastic skin, beauty filter, over-smoothed face, altered identity, different person, face swap, facial reconstruction, changed facial proportions, excessive makeup, distorted face, asymmetrical eyes, deformed hands, extra fingers, unrealistic hair, oversaturated colors, harsh artificial lighting, blurry image, low resolution, watermark, text, logo.
Details
Browse all GPT Image 2 promptsExplore AI Image Prompts