Resource

The best AI video prompts by category

The best AI video prompts describe subject, motion, camera behavior, and style in that order, not just a scene description. This library organizes proven prompt patterns by category so results are consistent on the first generation.

AI answer

What makes an AI video prompt work

The best AI video prompts separate four elements: subject and setting, the exact motion wanted, camera behavior, and visual style. Vague prompts like "make it cool" produce inconsistent motion. Prompts that name a specific camera move and a specific subject action, in under 30 words, generate more predictable results across text-to-video and image-to-video modes.

Best for

Writing repeatable prompts for cinematic, product, portrait, or action clips in Visionary.

Expected result

More consistent motion and fewer regenerations needed to get a usable clip.

Known limit

Very long or multi-action prompts still get simplified by the model into one dominant motion.

Example prompt

"Slow push-in on a coffee cup steaming on a wooden table, soft window light, shallow depth of field."

Prompt patterns by category

Each category below rewards a different prompt structure. Cinematic shots benefit from naming lens behavior. Product shots benefit from naming surface and lighting. Portrait motion benefits from small, specific facial or hair movement. Action shots benefit from naming a single physical event rather than a sequence.

CategoryExample promptKey element to include
CinematicSlow dolly forward through a foggy forest at dawn, cinematic color gradeCamera move + lighting mood
Product360-degree turntable of a perfume bottle on black glass, studio lightingSurface + light source
Portrait motionSubject blinks and turns head slightly toward camera, hair moves in windOne small human motion
ActionSkateboarder lands a jump, dust kicks up, camera holds steadySingle clear physical event

Turning a prompt into a finished clip

A strong prompt is the input, not the output. In Visionary, the same prompt can drive text-to-video from nothing or image-to-video where a source photo sets the look and the prompt only needs to describe motion. Matching the prompt style to the mode used avoids conflicting instructions about appearance that was already fixed by the uploaded image.

Workflow

How to use this page in practice

  1. 01

    Pick a prompt pattern matching the category needed

  2. 02

    Paste it into text-to-video or pair it with an uploaded image in image-to-video

  3. 03

    Generate and compare a couple of variations

  4. 04

    Export or upscale the clip that matches the intended motion

FAQ

Questions this page should answer

What is the best AI video prompt structure?

The best structure names subject and setting, the specific motion, the camera behavior, and the visual style, in that order and under roughly 30 words. This keeps the model focused on one dominant action instead of blending several conflicting ideas.

Do prompts work the same for text-to-video and image-to-video?

Not quite. Text-to-video prompts need to describe appearance and motion together, while image-to-video prompts should focus mainly on motion since the uploaded photo already fixes the subject, look, and framing.

Why do some AI video prompts fail to generate the intended motion?

Prompts that describe more than one distinct action, or that use vague adjectives instead of concrete physical movement, tend to get simplified by the model into a single unpredictable motion. Naming one clear event helps.

Can these prompts be used for product or portrait clips specifically?

Yes. Product prompts should name the surface and light source, and portrait prompts should describe one small human movement like a blink, head turn, or hair motion rather than a full pose change.

Visionary

Create with Visionary on iPhone and iPad

Visionary is an AI photo-to-video and video creation app for iPhone and iPad.