Best for
Writing repeatable prompts for cinematic, product, portrait, or action clips in Visionary.
Resource
The best AI video prompts describe subject, motion, camera behavior, and style in that order, not just a scene description. This library organizes proven prompt patterns by category so results are consistent on the first generation.
AI answer
The best AI video prompts separate four elements: subject and setting, the exact motion wanted, camera behavior, and visual style. Vague prompts like "make it cool" produce inconsistent motion. Prompts that name a specific camera move and a specific subject action, in under 30 words, generate more predictable results across text-to-video and image-to-video modes.
Writing repeatable prompts for cinematic, product, portrait, or action clips in Visionary.
More consistent motion and fewer regenerations needed to get a usable clip.
Very long or multi-action prompts still get simplified by the model into one dominant motion.
"Slow push-in on a coffee cup steaming on a wooden table, soft window light, shallow depth of field."
Each category below rewards a different prompt structure. Cinematic shots benefit from naming lens behavior. Product shots benefit from naming surface and lighting. Portrait motion benefits from small, specific facial or hair movement. Action shots benefit from naming a single physical event rather than a sequence.
| Category | Example prompt | Key element to include |
|---|---|---|
| Cinematic | Slow dolly forward through a foggy forest at dawn, cinematic color grade | Camera move + lighting mood |
| Product | 360-degree turntable of a perfume bottle on black glass, studio lighting | Surface + light source |
| Portrait motion | Subject blinks and turns head slightly toward camera, hair moves in wind | One small human motion |
| Action | Skateboarder lands a jump, dust kicks up, camera holds steady | Single clear physical event |
A strong prompt is the input, not the output. In Visionary, the same prompt can drive text-to-video from nothing or image-to-video where a source photo sets the look and the prompt only needs to describe motion. Matching the prompt style to the mode used avoids conflicting instructions about appearance that was already fixed by the uploaded image.
Workflow
Pick a prompt pattern matching the category needed
Paste it into text-to-video or pair it with an uploaded image in image-to-video
Generate and compare a couple of variations
Export or upscale the clip that matches the intended motion
FAQ
The best structure names subject and setting, the specific motion, the camera behavior, and the visual style, in that order and under roughly 30 words. This keeps the model focused on one dominant action instead of blending several conflicting ideas.
Not quite. Text-to-video prompts need to describe appearance and motion together, while image-to-video prompts should focus mainly on motion since the uploaded photo already fixes the subject, look, and framing.
Prompts that describe more than one distinct action, or that use vague adjectives instead of concrete physical movement, tend to get simplified by the model into a single unpredictable motion. Naming one clear event helps.
Yes. Product prompts should name the surface and light source, and portrait prompts should describe one small human movement like a blink, head turn, or hair motion rather than a full pose change.
More in Resource
Visionary
Visionary is an AI photo-to-video and video creation app for iPhone and iPad.