Start with a readable shot
Name the shot size, subject, location, and time of day. For image-to-video, treat the reference image as frame one and ask for motion that fits the existing pose.
Structure text-to-video and image-to-video shots, then test your prompt with a model enabled in this workspace.
AI video prompt workflow
A good AI video prompt describes a shot, not a pile of effects. Use Genzy to organize the subject, movement, camera, environment, and ending before testing the prompt with an enabled video model.
Prompt formula
Subject and starting frame + one primary action + one camera movement + environmental response + stability constraint + final state.
Name the shot size, subject, location, and time of day. For image-to-video, treat the reference image as frame one and ask for motion that fits the existing pose.
Use a slow push-in, lateral track, orbit, pull-back, or locked camera. One controlled camera instruction is usually easier to reproduce than several competing directions.
State that the face, product geometry, horizon, logo, or background structure should remain consistent. Stability instructions are especially useful for short commercial shots.
Copy and adapt
Replace the subject and setting, but keep the role of each sentence: composition, action or material, camera or light, protected details, and final use.
“Medium close-up. The same woman turns slightly toward the window as warm light reaches her face. Slow lateral camera track, subtle breathing and one natural blink. Preserve facial identity and room geometry. Settle on a stable three-quarter portrait.”
Why it works: One facial action, one camera move, and a defined ending reduce drift.
“Begin with the bottle in partial shadow. A soft highlight travels across the glass while the camera makes a slow push-in. Keep the bottle still and preserve its silhouette, cap, and label. End on a centered hero frame.”
Why it works: The light creates the reveal while protected product geometry stays fixed.
“Vertical 9:16 shot. The cyclist moves steadily toward the camera through light rain. The camera retreats at matching speed, keeping the rider centered. Preserve identity and bicycle frame; leave the upper quarter visually quiet for copy.”
Why it works: Mobile-safe positioning and a copy zone make the output easier to use.
“Wide shot from behind the hiker. The camera rises slowly from shoulder height to reveal the valley as morning fog moves between distant ridges. Keep the horizon level and the hiker anchored in the lower third. Finish on the full landscape.”
Why it works: Foreground anchor, restrained crane motion, and depth create a readable reveal.
Troubleshooting
Save the last successful output, change one instruction, and compare the result against a specific failure rather than rewriting the entire creative brief.
Replace appearance words with an observable action, direction, and stopping point.
State the starting position, travel direction, pace, and which subject remains centered.
Reduce rotation and duration, lock a few identity details, and test with a simpler camera move.
Shorten the action and ask the camera and subject to decelerate into a stable final frame.
Include the subject, one main action, one camera behavior, a small amount of environmental motion, and the condition the final frame should reach.
The source image already defines appearance and composition. The prompt should focus on what changes after that first frame and what visual details must stay stable.
No. The model menu reflects the providers enabled for the workspace. The prompt methods on this page are designed to remain useful across supported video models.
Use a sharp reference, limit head rotation and expression, protect a few identity traits, and test facial motion with a locked camera before adding camera movement.
Length matters less than structure. A few precise sentences covering action, camera, environment, stability, and ending are more useful than a long list of style adjectives.