Back to blog

How to Prompt AI Video Tools: 10 Techniques That Actually Work

Akshat Jain

Product & craft

PromptingCraftFebruary 05, 20265 min read

Get better results from AI video generators. 10 prompting techniques with examples and a template you can copy.

Most people prompt an AI video model the way they'd type a search query — a few nouns and a hope. Then they blame the model when the result is stiff, generic, or ignores half of what they asked. The models are better than that. The gap is almost always in the prompt. Here are the techniques that actually move the output, drawn from how these models genuinely behave.

Write a shot, not a search query.
01Section

1. Write a shot, not a search query

"woman walking city" is a search query. "A woman in a red coat walks briskly across a rain-slicked crosswalk at night, neon signs reflecting in the puddles" is a shot. The model has far more to work with in the second, and almost none of it is decoration — every clause is a decision it would otherwise make randomly.

02Section

2. Use plain sentences, not keyword salad

It is tempting to stack comma-separated tags: "cinematic, 8k, dramatic, moody, film grain." Current models read natural language better than tag soup. A plain descriptive sentence outperforms a pile of adjectives, because the model can parse relationships between words rather than guessing which of twelve tags matters most.

03

3. Separate subject, motion, and camera

The clearest prompts describe three things in order: what is in the shot, how it moves, and how the camera moves. Blur these together and the model has to disentangle them. Keep them distinct — subject, then subject motion, then camera language — and adherence jumps.

04

4. Describe motion, not the still

This is the single biggest fix for image-to-video. When you start from an image, the model can already see the picture — describing what is in the frame is wasted words. Spend the prompt on what happens: "the camera slowly pushes in as she turns her head toward the window." Describe the motion, not the photograph.

05

5. One camera move per beat

"The camera pans left, then cranes up, then zooms in" asks for three moves in a few seconds and usually produces a confused blur. Keep it to a single, clear camera move per generation. If you need a sequence, generate the beats separately and cut them together — you will get cleaner motion and more control.

06

6. Name real camera language

Models are trained on film, and they respond to film vocabulary: *dolly in*, *tracking shot*, *low angle*, *shallow depth of field*, *handheld*. These are not magic words, but they are precise instructions the model actually understands, where "make it look cinematic" is not.

07

7. Set the light and the atmosphere explicitly

Lighting is where "fine" becomes "striking." Say where the light comes from and what it does: *golden hour backlight*, *hard overhead light*, *soft window light from the left*. Add the atmosphere — *fog*, *dust in the air*, *volumetric light* — and the frame gains depth the model will not add on its own.

08

8. Match the technique to the model

Models have real, distinct strengths, and the same prompt does not behave identically across them. This is worth learning per model rather than assuming:

09

9. Use first and last frames when the model supports it

Several models let you set the opening and closing frame of a clip. This is the most underused control there is. Fixing where a shot begins and ends is the practical difference between clips that cut together and clips that drift — it turns generation from a slot machine into direction.

10

10. Iterate on one variable at a time

When a result is close but wrong, resist rewriting the whole prompt. Change one thing — the camera move, or the light, or the subject's action — and regenerate. You learn what the model is actually responding to, and you converge on the shot instead of thrashing between unrelated attempts.

11

What actually separates good prompts from bad

If there is a single principle underneath all ten, it is this:

Specific and singular — one clear subject, one motion, one camera move, described in a plain sentence, wins almost every time.

⚠️ Vague but hopeful — "cinematic, epic, beautiful" gives the model nothing to act on, so it falls back on the average of its training data. That average is what "generic AI video" looks like.

Overloaded — three camera moves, five adjectives, and two subjects in one prompt asks the model to resolve conflicts you could have resolved yourself.

12

Try it on a real shot

Prompting is a skill you build by watching what changes when you change the prompt — which is fast and cheap when you can pick the right model per shot and see the result in seconds. Describe one specific shot, generate it, then change exactly one thing and generate again. Two iterations in, you will prompt better than most people ever bother to.

That’s the piece.

Liked this piece? Share it.
Share
In this series
Keep reading

Try the craft

Describe one shot the way you’d brief a camera.

Open Ekly