You ask someone to water a plant, and the video slowly pushes into their face. The action is usable, but the shot no longer matches the wide frame you planned. Repeating “no zoom” may not solve it. First separate the camera request from the subject action, then find out what is actually moving.
Download the fixed camera landmark check · Plan a fixed-frame shot in Video Lab
Check whether it is really a zoom
A person getting larger does not prove that the camera zoomed. The person might walk toward a stationary camera. The background might also be reshaped by generation, making the frame appear to creep forward. These failures need different responses.
| Observed change | Possible explanation | Useful evidence |
|---|---|---|
| Background landmarks spread away from a common center | Zoom-like scaling or forward camera movement | Compare both nearby and distant fixed objects |
| Most of the scene slides in the same direction | Pan-like drift or a translated crop | Check several stationary edges |
| The person grows while the room stays fixed | Subject approaching the camera | Watch feet, position and foreground overlap |
| One shelf bends while the window stays fixed | Local geometric warping | Compare object shape, not only location |
| Only an object entering the frame changes size | Object motion or transformation | Track its outline against the static room |
These clues do not recover a real camera from generated pixels. They help you decide whether to simplify the action, change the camera brief or reject a malformed scene. Call the problem a zoom only after checking more than the main subject.
Define which movement belongs in the shot
Write two short lines before generating: what moves, and what stays fixed. For a plant-watering shot, the hand lifts a small watering can and tips it once. The bench, pot, shelves and camera stay in the same positions. Leaves can react gently without making the whole scene sway.
Remove a walking approach, a sweeping reveal or a focus pull if you do not need it. Each additional event gives the model another thing to interpret. A fixed shot can still contain useful action, but its energy should come from the action you actually requested.
Also inspect the source image. A dramatic close crop or strong motion blur may suggest a different visual continuation from a quiet wide shot. If the final video needs the whole bench, begin with the whole bench visible. A prompt cannot reliably recover the exact off-frame workshop you imagined.
Use positive camera language and one clear action
Runway’s Image to Video prompting guide separates subject action, environment and camera motion. It recommends starting simply and adding detail as needed. It also acknowledges that reducing generated camera movement may still require editing. A camera instruction is a request to the model, not a physical lock.
Adapt the objects to the actual source. Do not describe a window landmark that is absent from the image. For text-to-video, describe enough of the setting to establish the composition, then keep the motion request equally narrow.
Read the rest of the prompt for contradictions. Phrases such as following the hands, pushing closer, dramatic reveal or orbiting the subject ask for a different camera. A long list of prohibitions cannot be assumed to cancel those requests. Remove the contradictory instruction first.
If the selected model exposes a camera preset or negative-prompt field, use its documented controls. Syntax and available options vary. Do not paste model-specific camera tokens into another tool and assume they have the same effect.
Mark three stationary landmarks before the retry
Choose one landmark toward the left, one toward the right and one at a different depth if the scene provides it. For our proposed workshop exercise, use a window-frame corner, a shelf intersection and the base of the stationary plant pot. Avoid leaves, hands, screens or reflections that can legitimately move.
Capture the first, middle and final usable frames at the same dimensions. Place them side by side or on separate layers in an editor. Record each landmark’s approximate horizontal and vertical position, plus any change in the shape around it. A simple grid is enough for a visual review.
If you measure positions, express them consistently. For example, divide horizontal pixel position by frame width so a later resized export can be compared. The worksheet leaves these values blank because the useful tolerance depends on the delivery, crop and scene. There is no invented universal percentage that makes a shot stable.
Two moving landmarks and one fixed landmark deserve a closer look. They could reflect depth-dependent camera movement, but they could also reflect local warping. Examine straight edges and object proportions. If a shelf changes from four compartments to three, the problem is not merely a shaky view of the same shelf.
Change one cause at a time
Keep the source frame and action unchanged for the first camera-focused retry. Simplify the camera language and remove conflicting instructions. Save the request and output together so you can tell what changed. With a stochastic generator, one better result is a usable candidate, not proof that the wording always works.
If the camera stays steady but the person does not complete the action, simplify the action next. Try tipping a can already held above the plant instead of lifting, carrying, pouring and returning it. This reduces the work happening inside the same shot.
If the background warps, try a cleaner source with fewer repeated fine patterns or less visual ambiguity. Keep the room recognizable rather than redesigning everything at once. If only one short moment is usable, decide whether that moment can serve the edit before spending more generations on a longer shot.
Know what stabilization can and cannot rescue
A video editor may compensate for small global movement by shifting, scaling or cropping the frame. That consumes image area and can soften the export. Leave room around important objects if this is a likely finishing step, and inspect the actual output dimensions afterward.
Stabilization is less useful when the room changes shape independently. It cannot restore a shelf that melts, a hand that changes anatomy or an object that disappears. Trying to stabilize such a clip may make one region look steady while another stretches more visibly.
For a scene that truly needs a fixed background, another production approach is a static background plate with a separately composited moving element. That requires an editor, a suitable mask and matching light. It is a fallback workflow, not a promise that a single text prompt will create independently editable layers.
Avoid freezing the entire clip as the first repair. The frame may become stable, but the intended action is then gone. Decide whether a shorter cut, a restrained crop or a new source better preserves the purpose of the shot.
Accept the shot against its actual use
A subtle drift that is harmless in an atmospheric montage may break a split-screen comparison or a repeated product demonstration. Record the intended use before approving the clip. If it has to line up with another take, compare the relevant background edges in the final edit.
Watch once for framing, once for the action and once for geometry. Separating those passes prevents a pleasing gesture from distracting you from a bending window frame. Then watch the whole sequence at normal speed with the surrounding shots.
Keep the shortest version that communicates the action cleanly. A stable start does not excuse a late push-in that changes the composition during the important moment. Trim only if the result still has a natural beginning and end.
How QuestStudio helps
Open Video Lab and select an appropriate generation mode. Where the model supports image-to-video, supply the approved starting frame and the fixed-camera brief. Use the landmark worksheet to compare candidates; available model controls and durations vary.
Save the source, camera wording and accepted clip as one shot record. Begin with a single action and verify the background before expanding the sequence. The goal is a usable fixed-frame shot, not a long collection of increasingly forceful camera instructions.
The examples are proposed production exercises, not measured tool tests. The original hero image is an AI-generated editorial illustration.
Frequently asked questions
Why does my AI video keep zooming despite a no-zoom prompt?
The prompt or source may imply a moving shot, and the model may not obey a camera instruction precisely. Remove conflicting language, describe a stationary camera positively and inspect a short retry.
Does locked-off guarantee a motionless camera?
No. It states the intended shot. Verify several stationary landmarks across the generated clip before accepting it.
How can I tell camera motion from a person walking forward?
Compare background landmarks. If they stay fixed while the person grows and changes position, the movement may belong to the subject.
Can stabilization fix warped scenery?
It may correct small global movement, usually with a crop. It cannot reliably repair local objects that change shape or disappear.
Should I use the same camera syntax in every model?
No. Use documented controls for the selected model. A camera token, preset or negative-prompt field may not exist or behave the same way elsewhere.

