
10 Cinematic Camera Shot Prompts for GPT Image 2
10 copy-ready cinematic camera shot prompts for GPT Image 2, plus a practical shot formula, continuity workflow, and composition fixes.
Putting cinematic near the end of an image prompt often produces a polished frame, but not a deliberate one: the subject lands somewhere near the middle, the camera feels undecided, and every variation tells a different story.
The fix was not a longer style list. I started naming the job of the shot first. An establishing shot tells us where we are. A close-up asks us to read a face. A low angle lends weight. Once that decision is explicit, lighting and color stop carrying the whole prompt.
This guide turns ten useful camera shots into copy-ready GPT Image 2 prompts. They all stay inside one original mountain-village story world, so you can see how framing changes the emphasis while the character, setting, and visual continuity remain recognizable.

The prompt formula I use
[SUBJECT + ACTION] in [SETTING].
[SHOT SIZE], [CAMERA ANGLE], [LENS LOOK].
[LIGHT DIRECTION], [COLOR PALETTE], [DEPTH OF FIELD].
[CONTINUITY DETAILS], [EXCLUSIONS].Use one continuity lock above all ten prompts:
The same adult male traveler with dark wavy hair, a plain black jacket,
black backpack, remote mountain village at sunrise, warm light from camera-left,
cool blue shadows, grounded photorealism, no branding.The order matters. Put the subject and shot before the aesthetic finish. If the camera instruction arrives after a paragraph of mood words, it is easier for the composition to drift.
One correction to a lot of prompt cards online: 35mm, f/2.8, and ISO 200 are not real capture settings inside an image model. They are visual cues. If you need the result rather than the camera cosplay, write the result too: moderate wide-angle perspective, shallow depth of field, or clean low-noise shadows.
1. Establishing shot: orient the viewer
Use this for the first frame of a sequence, a landscape-led image, or any scene where place matters more than facial detail.
A lone traveler arrives at a remote mountain village at sunrise.
Wide establishing shot from a distant hillside, 24mm wide-angle look,
the traveler small in frame, the road leading toward the village.
Warm dawn light from camera-left, cool blue shadows, layered mist,
deep depth of field, restrained film grain, no text, no logos.2. Full shot: show body language
A full shot keeps the person readable without losing the environment. It is useful for fashion, action poses, and character introductions.
The same traveler pauses on a stone path and tightens a backpack strap.
Full-body shot at eye level, 35mm natural perspective, both feet visible,
village houses framing the path, warm dawn rim light, realistic fabric,
moderate depth of field, same black jacket and backpack, no extra people.3. Medium shot: balance person and place
This is the reliable conversational frame: roughly waist-up, enough face to read, enough background to preserve context.
The traveler turns after hearing a bell behind him.
Medium waist-up shot at eye level, 50mm lens look, subject on the right third,
soft village background, warm side light, natural skin texture,
subtle concern in his expression, same wardrobe, no dramatic motion blur.4. Close-up: make the emotion the scene
Do not ask a close-up to explain the whole location. Give the background one or two quiet cues and let the face do the work.
The traveler studies a hand-drawn map, realizing the marked bridge is gone.
Tight facial close-up, eye-level camera, 85mm portrait look,
eyes sharp, map edge softly visible at the bottom of frame,
warm dawn highlight on one cheek, shallow depth of field,
natural pores and fine hair, no beauty retouching, no text.5. Extreme close-up: isolate one clue
An extreme close-up works when one detail carries the beat: an eye, a hand, a cracked surface, a mechanical switch.
Extreme close-up of the traveler's eye as he looks toward a broken rope bridge.
100mm macro look, eyebrow and skin texture filling the frame,
the bridge only a small soft corneal reflection, not a crisp miniature scene,
narrow focus plane,
soft amber catchlight, controlled contrast, no duplicated iris,
no surreal anatomy, no text.6. Low-angle shot: add weight
Low angles can make a character feel capable, threatening, or simply larger than the setting. Keep vertical lines intentional so the frame does not look accidentally warped.
The traveler stands on a ridge above the village as wind catches his coat.
Low-angle full shot from knee height, 24mm dramatic perspective,
figure centered against a wide sky, stone wall anchoring the foreground,
hard dawn rim light, deep slate and amber palette,
straight architectural lines, natural hands, no logos.7. High-angle shot: reduce the subject
High angle is not automatically top-down. The camera is above the subject, but the horizon or surrounding walls can still be visible.
The traveler stands alone in a small village courtyard, choosing between two paths.
High-angle shot from a second-story balcony, 35mm look,
subject below center, paths forming a clear Y shape,
long morning shadows, quiet cool palette with warm window light,
deep focus, no crowd, no signs, no text.8. Bird's-eye view: turn the scene into geometry
For a true overhead result, say camera pointing straight down. That removes the ambiguity hidden inside “aerial.”
The traveler crosses a patterned stone square while market tables open around him.
True bird's-eye view, camera pointing straight down, 24mm overhead look,
subject small but readable, tables and paths forming clean geometry,
early sunlight creating long parallel shadows, deep focus,
no oblique perspective, frame edges aligned with the square,
no drone visible, no text.9. Over-the-shoulder shot: give the viewer a position
Over-the-shoulder framing is often more useful for still images than a vague “follow shot.” It tells GPT Image 2 whose perspective matters and what they are looking at.
Over the traveler's right shoulder as he compares the map with the village gate.
Foreground shoulder and map softly out of focus, gate sharp in the middle distance,
50mm lens look, eye-level camera, layered composition,
warm morning light, same black jacket and paper map,
only two hands visible, no readable writing, no logos.10. POV shot: put the viewer inside the action
POV works best when hands or an object prove where the camera is. Without that anchor, many results collapse back into an ordinary wide shot.
First-person POV walking across a narrow wooden bridge toward the mountain village.
Both gloved hands holding the bridge ropes at the lower edges of frame,
natural 24mm perspective, planks leading forward, slight height above the river,
warm sunrise ahead, realistic hand anatomy, sharp path with softer periphery,
no visible face, no third-person body, no text or logos.Shot size and camera movement are different controls
The source list that inspired this test mixed shot sizes with movements such as a follow shot and a crane move. That is normal in filmmaking, but it matters when prompting.
- Shot size says how much of the subject we see: wide, medium, close-up.
- Angle says where the camera is: low, high, overhead, POV.
- Movement says what the camera does over time: follow, crane up, dolly in, orbit.
For a still image, translate movement into a frozen visual result: rear three-quarter frame with directional background blur or elevated reveal from above the rooftops. For video, write the move and its timing directly.

A three-pass workflow that wastes fewer generations
- Lock the story facts. Keep the same character, wardrobe, location, time of day, and color palette in every prompt.
- Change one camera slot. Compare establishing versus medium before changing wardrobe, weather, and lighting too.
- Name the failure, not a new vibe. Replace
make it more cinematicwithmove the camera lower,leave more negative space, orkeep both hands in frame.
If you want to build your own version, start in the GPT Image 2 generator, then use the broader prompt-writing guide to tighten subject, style, text, and composition. For designed layouts rather than film frames, the poster prompt collection uses a separate eight-block structure.
Bottom line
“Cinematic” is a finish. The shot is the decision.
Pick the narrative job first, describe the visible result of the lens language, and keep continuity facts fixed while you test. Ten camera choices are enough to turn one plain scene into a sequence that feels directed rather than randomly decorated.
Weitere Beiträge

Was ist GPT Image 2? Eine vollständige Einführung
GPT Image 2 ist OpenAIs neues, multimodales Bildmodell — das erste, das nicht-lateinischen Text und komplexe Layouts zuverlässig beherrscht. Alles, was du wissen musst.

8 GPT Image 2 Design Prompts for Real-World Work
Copyable GPT Image 2 design prompts for worldbuilding, diagrams, storyboards, picture books, product ads and packaging—with 4 generated examples.

GPT Image 2 vs Nano Banana 2 vs Midjourney v7 (2026)
GPT Image 2 vs Nano Banana 2 vs Midjourney v7 — welches KI-Bildmodell gewinnt bei Text, Postern, Fotos und Concept Art? Ein praktischer Entscheidungsleitfaden für 2026.
Generate your first image with GPT Image 2 — right now
Reliable non-Latin text rendering, directed editing, and 50+ ready-to-use prompts. No downloads — just open in your browser.