Loop Forge
Loop ForgeMiniMax H3 camera shots

Crash zoom in

zoom
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still on a snow-covered arctic coastline, as the camera holds a locked-off wide shot and then performs a single violent crash zoom in to a tight close-up.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke Varotal 18-100mm zoom lens, locked off on a tripod, cold blue-green arctic lighting with the aurora glowing softly overhead, a slightly desaturated palette.
[Shot 1] The shot opens wide, with <Subject 1> standing small and full-figure on the snowy rocky shore, with dark rocky cliffs, patches of green aurora borealis in the sky, still water, and an old wooden sailing ship with tall masts anchored behind her. The camera is locked off and completely motionless, and the framing does not change at all for the first two seconds. At the two-second mark the camera performs a Zoom In with large amplitude at fast speed: a single violent crash zoom that travels the whole distance from the wide shot to a tight close-up of <Subject 1>'s face in about 200 milliseconds, in one continuous snap, with no easing at either end and nothing gradual about it. The instant the zoom lands the camera is locked off again in the tight close-up and the framing does not change for the remainder of the shot. <Subject 1> stands completely still with a blank, neutral expression for the whole of the opening wide shot. Her reaction begins only after the zoom has landed: her eyes widen and focus, her brow lifts and tightens, her lips part slightly, and she settles into a held, wide-eyed stillness for the rest of the shot.

overall_soundscape: A faint cold wind and distant creaking from the ship's rigging hold steady and completely unchanged through the opening two seconds. At the two-second mark a single hard impact lands in one frame as the zoom arrives, and the ambience cuts instantly from open and distant to tight, close and intimate, holding there to the end. Nothing builds gradually at any point.

non_diegetic_music: Silence through the opening. One sharp percussive stab on the single frame the zoom lands, then silence again.

Yo-yo zoom

zoom
frames
192
length
8.00s
size
1344×768
Frame from the yo-yo zoom shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long dark brown wavy hair, warm olive skin and a bright open smile, wearing a white lace-trimmed cropped camisole and high-waisted beige trousers. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a sunlit city rooftop terrace at golden hour, in a tight close-up, as the camera performs a yoyo zoom.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke Varotal 18-100mm zoom lens, locked off on a tripod, warm golden-hour sunlight with strong lens flare, a rich saturated palette.
[Shot 1] A tight close-up fills the frame with <Subject 1>'s face. She is standing in place on the warm rooftop deck, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with low sun flaring between distant towers, potted plants and strings of lights behind her. The camera begins in a tight close-up on <Subject 1>'s face, then zooms out, starting slowly and accelerating harder and harder as it goes until the whole view smears into motion blur and she has fallen away into the far distance, a tiny figure on the rooftop. The camera holds there for about a second, then zooms back in with large amplitude at fast speed, snapping across the entire distance in a single instant to land tight on her face again exactly as it began. <Subject 1> holds her ground throughout, breathing, blinking and letting her hair move in the rooftop breeze, her expression shifting from an easy smile into startled bewilderment as the world tears away from her and slams back.

overall_soundscape: Rooftop ambience close and intimate, stretching thin and far away as she recedes, then slamming close and present again in a single instant.

non_diegetic_music: A tone that bends downward and stretches as everything races away from her, then snaps back up to pitch in one beat.

Dolly zoom

specialty
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still centre stage on a concert stage with tall deep-red velvet curtains, a single warm spotlight beam falling from high above, a chrome microphone on a tall stand, a drum kit on a riser behind, stacked keyboards to the left, monitor wedges, and a polished wooden stage floor reflecting the warm light, in a waist-up medium close-up, as the camera performs a dolly zoom.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke S4 lens at 50mm, warm amber stage lighting with a hard spotlight from above and deep shadow beyond it, rich saturated reds.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing completely still centre stage, lit by the overhead spotlight, with the microphone stand just in front of her and the red curtains, drum kit and keyboards behind her. The camera pushes in while simultaneously zooming out on the Cooke S4 lens, so that <Subject 1>'s size and position in frame stay exactly constant throughout, while the red curtains, the drum kit, the keyboards and the microphone stand behind her recede, shrink, and spread apart into much greater depth as the lens's focal length changes. As the shot plays out, <Subject 1>'s expression shifts from blank and neutral into dawning realization and slight shock: her eyes widen and focus more intensely, her brow lifts and tightens, her lips part slightly, and her breathing grows visibly shallower and faster, the change building through the middle of the shot and settling into a held, wide-eyed stillness by the end.

overall_soundscape: The quiet room tone of a large empty auditorium continues throughout, with faint distant creaks and <Subject 1>'s breathing growing faster and shallower as the shot progresses.

non_diegetic_music: A single sustained low string note swells partway through the shot, then fades toward the end.

Snorricam

specialty
frames
192
length
8.00s
size
864×480
Frame from the snorricam shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long light brown wavy hair, fair freckled skin, round gold-rimmed glasses, wearing a simple white slip dress with thin straps. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> running through a crowded indoor house party from room to room, in a head-and-shoulders close-up from just below her eyeline, as the camera performs a snorricam shot.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa Mini with a 14mm wide prime, the camera fixed rigidly to her torso so it moves exactly with her body, harsh mixed party lighting with warm tungsten sconces and coloured LED wash spilling from open doorways sweeping hard across her face as she passes each one, a rich saturated palette with deep shadow.
[Shot 1] A head-and-shoulders close-up of <Subject 1>, framed from just below her eyeline a short distance in front of her and angled gently upward at her face, shows her whole head and both bare shoulders with a clear margin of open space above her hair, so that the top of her head never reaches the top of the frame, as she is running hard down a narrow, packed indoor corridor at a house party and banking round its corners into the rooms off it, her hair flying and the thin straps of her dress shifting with every stride, with people crushed shoulder to shoulder along both walls with drinks, shouting over the music, open doorways spilling coloured light ahead of her, a low ceiling strung with warm bulbs close overhead, and each room she banks into more congested than the last. <Subject 1> runs hard down the corridor and banks hard round each corner and through each doorway into the next crowded room, carried forward at the same speed from the first frame to the last, never breaking stride. Her shoulders and chest stay frozen in the frame throughout: they hold exactly the same place at exactly the same size in every single frame, never drifting a pixel, and her head is the only part of her that moves inside that frozen frame. Everything else is thrown onto the world. Tracking with her at fast speed and shaking strongly with large amplitude, the whole party - the ceiling lights, the door frames, the packed bodies pressed in on both sides - heaves and drops in exact time with her stride, and each time she banks into a turn the entire room whips sideways across the frame in one fast pan, the walls and the crowd sweeping past the lens fast enough to smear into streaks before the next room opens up ahead of her. <Subject 1> is already wide-eyed and frightened in the very first frame, her brows pulled up and together and her eyes stretched wide, and the fear only tightens as she runs: her stare fixes harder, her brows draw in further, and her breathing grows faster and more ragged through her nose. Her lips stay pressed together the whole way, the breath forced out through her nose. Her head rocks with the impact of every stride while her shoulders stay locked in place, and both arms swing freely and loosely at her sides in rhythm with her run, her hands open and relaxed.

overall_soundscape: Her footfalls and breathing stay close, loud and constant, right up against the microphone, her breath tightening into a fast ragged hiss through her nose as the shot goes on, while the party swings and lurches around her with every stride, the music and shouting swelling hard as she banks into each new room and sweeping away behind her.

non_diegetic_music: A relentless two-note pulse locked exactly to her stride, one hit per footfall, never speeding up or slowing down, with a high thin dissonance creeping in above it and tightening as the shot goes on.

Rack focus

specialty
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long dark brown wavy hair, warm olive skin and a bright open smile, wearing a white lace-trimmed cropped camisole and high-waisted beige trousers. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.
<Subject 2> is the second young woman in <Picture 2>: long light brown wavy hair, fair freckled skin, round gold-rimmed glasses, wearing a simple white slip dress with thin straps. Her exact facial structure, features and proportions stay identical to <Picture 2> in every frame.
<Subject 3> is the third young woman in <Picture 3>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 3> in every frame.

summary:
[reference generation] The target video shows <Subject 1> in the foreground with <Subject 2> and <Subject 3> talking together far behind her in a sunlit city rooftop terrace at golden hour, in a medium-wide shot, as the camera performs a rack focus.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.
<Subject 2> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 2> in every frame with zero drift.
<Subject 3> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 3> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke S4 75mm prime wide open, locked off, focus pulled with a follow focus, warm golden-hour sunlight with strong lens flare, a rich saturated palette.
[Shot 1] A medium-wide shot holds two distinct depth planes at once: <Subject 1> close to camera and the rooftop opening out well behind her. She is standing in place on the warm rooftop deck, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with low sun flaring between distant towers, potted plants and strings of lights behind her. The camera holds a static shot. <Subject 1> stands close to camera in sharp focus while <Subject 2> and <Subject 3> stand talking together far behind her, soft and unfocused. <Subject 1> slowly turns her head to look back at them, and the focus racks from her face to the two women - she falls into soft blur exactly as they resolve sharply. <Subject 1>'s idle curiosity sharpens into recognition as she turns, while <Subject 2> and <Subject 3> stay absorbed in their own conversation, laughing quietly together.

overall_soundscape: The two women's quiet conversation and laughter sits distant and indistinct behind the rooftop ambience, then lifts forward and becomes clearly audible as the focus reaches them.

non_diegetic_music: A faint sustained pad that shifts harmony once, at the moment the focus changes.

Split screen

specialty
frames
192
length
8.00s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> seated at a desk in a dim room at night, in a three-panel split screen, as the camera performs a three-panel split screen.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa Mini with a 35mm prime, locked off, cool monitor light on her face against a warm dim room, a muted contemporary palette.
[Shot 1] The frame is divided into three equal vertical panels side by side, separated by thin black gutters. The left panel holds <Subject 1>, seated at a cluttered desk typing on a keyboard, a bright monitor in front of her, a mug and scattered papers beside it, a dark room and a rain-streaked window behind her. Every panel is locked off and completely static; nothing pans, zooms or moves at any point. The three panels open one after another at even intervals, two full seconds apart. The left panel plays from the very first frame, showing her from her left side in profile as she types, while the centre and right panels are solid black. The centre panel stays black for two whole seconds and comes on two seconds into the shot, showing the same moment from the front, framed over the back of the monitor with the screen glow on her face. The right panel stays solid black for two whole seconds longer than the centre one and comes on four seconds into the shot, showing her from behind in full, her back and the bright monitor in view. For the last four seconds all three panels play together in perfect sync, the same continuous typing at the same instant from three different angles. <Subject 1> types steadily and continuously throughout, her hands staying on the keyboard, pausing once to read the screen before carrying on.

overall_soundscape: Steady keyboard clatter, close and continuous, over the low hum of a quiet room, with a chair creaking once.

non_diegetic_music: A tight rhythmic pulse locked to the typing, a new layer entering each time another panel opens.

Whip pan

pan and roll
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.
<Subject 2> is the second young woman in <Picture 2>: shoulder-length dark brown wavy hair, fair skin and blue-grey eyes, wearing a chunky oatmeal cable-knit wool sweater. Her exact facial structure, features and proportions stay identical to <Picture 2> in every frame.

summary:
[reference generation] The target video shows <Subject 1> and <Subject 2> standing apart on a snow-covered arctic coastline, in waist-up medium close-ups, as the camera performs a whip pan.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.
<Subject 2> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 2> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa Mini with a Zeiss Ultra Prime 35mm, on a fluid-head tripod, cold blue-green arctic lighting with the aurora glowing softly overhead, a slightly desaturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing completely still on the snowy rocky shore, with dark rocky cliffs, patches of green aurora borealis in the sky, still water, and an old wooden sailing ship with tall masts anchored behind her. The camera starts framed on <Subject 1>, then whip pans right away from her and lands on <Subject 2> standing further along the shore, the cliffs and snow between them smearing into streaked horizontal motion blur through the middle of the move, the frame settling and resolving sharply on <Subject 2>'s face. <Subject 1>'s expression is caught mid-realization in the instant before the camera leaves her, and <Subject 2> is already turning to look back as the frame settles on her - the shot ends held on <Subject 2>, not on <Subject 1>.

overall_soundscape: A faint cold wind continues throughout, with distant creaking from the ship's rigging, and <Subject 1>'s breathing growing faster and shallower as the shot progresses.

non_diegetic_music: A single sustained low string note swells partway through the shot, then fades toward the end.

Dutch angle

pan and roll
frames
124
length
5.17s
size
1344×768
Frame from the dutch angle shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: shoulder-length dark brown wavy hair, fair skin and blue-grey eyes, wearing a chunky oatmeal cable-knit wool sweater. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a snow-covered pine forest at dusk, in a waist-up medium close-up, as the camera performs a dutch angle shot.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke S4 35mm prime, on a tripod with the head canted over, cold blue dusk light filtering down through the pines, a muted desaturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing completely still on the snow between the trees, with tall dark conifers receding into blue mist, snow lying heavy on the branches behind her. The camera rolls counterclockwise into a canted dutch angle and holds there, the horizon tilted, <Subject 1> off-axis. As the shot plays out, <Subject 1>'s expression shifts from blank and neutral into dawning realization and slight shock: her eyes widen and focus more intensely, her brow lifts and tightens, her lips part slightly, and her breathing grows visibly shallower and faster, the change building through the middle of the shot and settling into a held, wide-eyed stillness by the end.

overall_soundscape: Deep forest hush, snow settling from branches, a distant crack of timber, and <Subject 1>'s breathing growing faster and shallower as the shot progresses.

non_diegetic_music: A single sustained low string note swells partway through the shot, then fades toward the end.

Super dolly in

tracking
frames
124
length
5.17s
size
1344×768
Frame from the super dolly in shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: shoulder-length dark brown wavy hair, fair skin and blue-grey eyes, wearing a chunky oatmeal cable-knit wool sweater. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a snow-covered pine forest at dusk, in an extreme wide shot, as the camera performs a super dolly in.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke S4 35mm prime, on a dolly running on track, cold blue dusk light filtering down through the pines, a muted desaturated palette.
[Shot 1] An extreme wide shot looks straight down a forest track, the trunks of two near pines rising close at the left and right edges of frame, and places <Subject 1> small and distant at the far end of it, standing in place on the snow between the trees, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with tall dark conifers receding into blue mist, snow lying heavy on the branches behind her. The camera pushes in with large amplitude at fast speed, charging the entire length of the track toward her: the two near pines sweep outward past the edges of frame and out of shot as the camera passes them, further trunks rush by on both sides and the snow streams underneath, and the move ends in a tight close-up with her face filling the frame. <Subject 1>'s expression hardens into defiance as the camera closes on her - jaw setting, chin lifting, her gaze steady and level by the end.

overall_soundscape: Deep forest hush and snow settling from branches, building into a rushing wall of air with snow and needles streaming past, then dropping away to close, still quiet on her face.

non_diegetic_music: A low sustained bass tone climbing steadily in pitch and volume through the whole shot, cut off dead on the final frame.

Eyes in

tracking
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still on a snow-covered arctic coastline, in a waist-up medium close-up, as the camera performs a slow push in to the eyes.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa with a Cooke S4 75mm prime, on a slider for a slow push, cold blue-green arctic lighting with the aurora glowing softly overhead, a slightly desaturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing in place on the snowy rocky shore, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with dark rocky cliffs, patches of green aurora borealis in the sky, still water, and an old wooden sailing ship with tall masts anchored behind her. The camera pushes in with large amplitude at slow speed toward <Subject 1>'s face and keeps going past it until her right eye fills the entire frame edge to edge, the iris and pupil filling the view. <Subject 1>'s expression tightens slowly into dread as the camera closes - pupils contracting, a single blink, then held wide and unmoving.

overall_soundscape: Cold wind and distant surf receding steadily into muffled quiet, until nothing is left but one slow heartbeat as the eye fills the frame.

non_diegetic_music: A high sustained string that tightens and narrows as the shot closes in, thinning to a single thread.

Aerial pullback

tracking
frames
124
length
5.17s
size
1344×768
Frame from the aerial pullback shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: shoulder-length dark brown wavy hair, fair skin and blue-grey eyes, wearing a chunky oatmeal cable-knit wool sweater. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a snow-covered pine forest at dusk, in a waist-up medium close-up, as the camera performs an aerial pullback.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on a RED V-Raptor with a 14mm wide prime, flown on a drone, cold blue dusk light filtering down through the pines, a muted desaturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing in place on the snow between the trees, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with tall dark conifers receding into blue mist, snow lying heavy on the branches behind her. The camera pulls out with large amplitude at fast speed while pedestalling up, rising and retreating until <Subject 1> is a small lone figure far below and the whole forest opens out around her. <Subject 1> stands alone and unmoving as she recedes, head tipping slowly back to take in the scale of the place as it opens around her.

overall_soundscape: Deep forest hush, snow settling from branches, a distant crack of timber, and <Subject 1>'s breathing stays audible throughout.

non_diegetic_music: A single sustained low string note swells partway through the shot, then fades toward the end.

Handheld

tracking
frames
124
length
5.17s
size
1344×768
Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long wavy ginger-red hair, fair skin with light freckles, wearing a cream short-sleeve t-shirt, a dark cord necklace with a metal cross pendant, high-waisted blue jeans with a brown leather belt and a small brown pouch, and woven cord wrist wraps on both wrists. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still on a snow-covered arctic coastline, in a waist-up medium close-up, as the camera performs a handheld shot.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa Mini with a Canon K-35 35mm, handheld and running with her, cold blue-green arctic lighting with the aurora glowing softly overhead, a slightly desaturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> running on the snowy rocky shore, hair flying and clothing snapping with the movement, with dark rocky cliffs, patches of green aurora borealis in the sky, still water, and an old wooden sailing ship with tall masts anchored behind her. The camera shakes strongly at fast speed while performing a tracking shot alongside <Subject 1> as she runs, the frame lurching and correcting with every stride. <Subject 1> runs with urgent purpose, lips pressed closed and jaw set, glancing back once over her shoulder without breaking stride.

overall_soundscape: A faint cold wind continues throughout, with distant creaking from the ship's rigging, and <Subject 1>'s breathing stays audible throughout.

non_diegetic_music: A single sustained low string note swells partway through the shot, then fades toward the end.

360 orbit

tracking
frames
124
length
5.17s
size
1344×768
Frame from the 360 orbit shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long dark brown wavy hair, warm olive skin and a bright open smile, wearing a white lace-trimmed cropped camisole and high-waisted beige trousers. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a sunlit city rooftop terrace at golden hour, in a waist-up medium close-up, as the camera performs a 360 orbit.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on an ARRI Alexa Mini with a Zeiss Ultra Prime 35mm, on a gimbal circling her, warm golden-hour sunlight with strong lens flare, a rich saturated palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> standing in place on the warm rooftop deck, breathing softly, hair and clothing stirring in the air, blinking and shifting her weight a little as people naturally do, with low sun flaring between distant towers, potted plants and strings of lights behind her. The camera performs an arc shot around <Subject 1> with large amplitude at fast speed, sweeping a complete circle around her and coming back to the front. <Subject 1> stays where she is through the sweep, turning her head slightly to keep the camera in view, a slow private smile spreading, hair lifting in the rooftop breeze as though she is the only calm thing in a turning world.

overall_soundscape: City traffic far below turning steadily around her, one continuous sweep of wind, strings of lights ticking past in turn.

non_diegetic_music: A circling arpeggio that completes exactly one full turn and lands back on the note it started on.

Crane rise

tracking
frames
124
length
5.17s
size
1344×768
Frame from the crane rise shot

Reference plates

Prompt
subject_definitions:
<Subject 1> is the young woman in <Picture 1>: long light brown wavy hair, fair freckled skin, round gold-rimmed glasses, wearing a simple white slip dress with thin straps. Her exact facial structure, features and proportions stay identical to <Picture 1> in every frame.

summary:
[reference generation] The target video shows <Subject 1> standing still in a sunlit summer meadow, in a waist-up medium close-up, as the camera performs a crane shot over the head.

retention_analysis:
<Subject 1> (appears in [Shot 1]): fully_preserved - her face, identity and clothing are held identical to <Picture 1> in every frame with zero drift.

detailed_description:
The target video is live-action and cinematic, shot on a RED V-Raptor with a 35mm prime, flown on a drone tracking alongside her, warm soft afternoon sunlight, a bright natural palette.
[Shot 1] A waist-up medium close-up frames <Subject 1> running steadily through the tall grass, hair flying and clothing snapping with the movement, with dry grass and wildflowers moving in the breeze, a soft distant treeline behind her. The camera starts low at <Subject 1>'s feet as she runs, then performs a tracking shot with large amplitude at fast speed, pedestalling up and tilting down to follow her from high overhead while she stays centred and sharp in frame. <Subject 1> runs at a steady rhythm with her lips closed and her jaw relaxed, chin level and hair flying, her expression easing out of hard concentration into open delight as the camera lifts away from her.

overall_soundscape: Her footfalls and the hiss of grass close by at the start, thinning out into open air, high wind and distant birdsong.

non_diegetic_music: A single held note that opens outward into a wide major chord across the shot.