Prompts
Veo 3.1 Prompt Guide
Copy-ready Veo 3.1 prompts for muvi.video. Covers prompt anatomy, 10 real examples by use case, when to reach for Seedance or Kling, and the iteration loop.
What You Need to Know First
Veo 3.1 is Google DeepMind's latest video generation model and it is available on muvi.video today across four variants, Fast, Quality, and two Image-to-Video options (ECL Image-to-Video and ECL Start/End). Every clip runs at 4, 6, or 8 seconds per generation, with output at 1080p (1920×1080 landscape or 1080×1920 portrait) in 16:9 or 9:16. Veo 3.1 is the model in the muvi.video catalog that most consistently delivers photorealistic commercial output: product textures render cleanly, native ambient audio (a product claim per Google's Veo documentation) matches scene context, and 1080p hero shots hold detail across the full frame. Veo 3.1 does not support deterministic seed control, iteration works by refining your prompt text across successive generations. See the pricing page for current Veo 3.1 coin costs and what each subscription tier includes.
The Seven Building Blocks of a Veo 3.1 Prompt
Veo 3.1 responds best to concrete, declarative scene descriptions. Adjective piles and abstract mood words underperform. These seven components, used in sequence, produce consistent 1080p output across commercial and cinematic formats.
1. Subject Name who or what occupies the primary focal point of the frame. Specificity at the subject level is where Veo 3.1 earns its commercial reputation: "a matte black ceramic coffee mug with a brushed steel lid" generates a specific object; "a coffee mug" generates something generic. For product work, treat the subject description as a product brief, the more detail you give here, the less guesswork the model applies to your hero object.
2. Environment Define the spatial context before you specify any movement or camera instruction. "a white marble surface inside a minimal photography studio, single window to the left casting soft daylight" anchors the model's spatial reasoning. Veo 3.1 holds environment consistency reliably within an 8-second clip when the spatial anchor is explicit. Omit the environment and the model will construct one, sometimes inconsistently with your lighting specification.
3. Action Write the action as a present-tense continuous statement. For product clips: "rotating slowly on its own axis, 180 degrees left to right." For scene clips: "a barista pours steamed milk from a steel pitcher into a cup." One primary action per prompt. Veo 3.1 handles layered simultaneous actions poorly, specify the most important one and let the model fill supporting motion.
4. Camera Language Veo 3.1 accepts standard cinematography vocabulary and executes it accurately: slow push-in, locked-off medium close-up, overhead flat lay, low angle establishing, rack focus from foreground to background. Include exactly one camera move or position instruction per prompt. Specifying two camera moves, "push in while also tilting up", produces unpredictable motion in most cases.
5. Lighting This is Veo 3.1's clearest strength for commercial work. Specify quality, direction, and color temperature together: "soft diffused studio light, 5600K, single overhead softbox with a white reflector filling the shadow side." Veo 3.1 reproduces specified lighting conditions with higher fidelity than most models in the catalog. Skipping the lighting instruction on a product shot hands the model your most important quality lever.
6. Audio Cue Veo 3.1 generates native ambient audio that matches scene context (a product claim per Google's Veo documentation). This runs automatically, a coffee shop scene will generate café ambient sound without any instruction. What you control with an explicit audio cue is specificity: "light rain against a window, no music, no dialogue" gets you a specific soundscape. If you are generating a scene where ambient audio would be distracting, such as a silent product rotation, write "no audio" explicitly. Do not use dialogue audio cues expecting lip-synced speech; Veo 3.1's audio strength is ambient and environmental, not dialogue-to-lip-movement matching.
7. Format and Iteration Anchor End each prompt with the target format: "Generate as an 8-second 16:9 clip at 1080p." Veo 3.1 does not support deterministic seeds, you cannot reuse a seed to lock a composition. Iteration instead works through prompt refinement: change one specific phrase in your prompt and regenerate. You will get a related but non-identical output that trends toward your new instruction. For tighter frame-level consistency across takes, use the Veo 3.1 ECL Image-to-Video variant with a reference still from a prior generation as your anchor.
Copy-Ready Veo 3.1 Prompts by Use Case
Each prompt below is labeled with the muvi.video Veo 3.1 variant it targets. Copy the prompt into muvi.video Studio, select the matching variant, and generate.
Product Hero Shot, Packaged Consumer Good
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A matte white skincare serum bottle with a gold dropper cap sits on a polished black marble surface. The bottle rotates slowly 180 degrees left to right. Overhead softbox at 5600K. A single thin highlight traces the gold cap edge as it rotates. No human subjects. No dialogue. Quiet studio ambient, almost silent. 8 seconds, 16:9 landscape, 1080p.
Product Hero Shot, Electronics
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A pair of matte black wireless over-ear headphones lies on a dark grey concrete surface. The camera performs a slow push-in from a wide shot to a medium close-up over the full 8 seconds, coming to rest on the right ear cup. Soft diffused top light, 5000K. No reflections on the headphone surface. No human subjects. No audio. 8 seconds, 16:9 landscape, 1080p.
Commercial B-Roll, Coffee Brand
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A barista's hands pour steamed milk from a steel pitcher into a ceramic cup on a wooden café counter. The camera holds a medium close-up locked off slightly above table height. Warm 3000K ambient café light from behind the counter. Steam rises from the cup. Background café ambient sound, low murmur, no music, no dialogue. 8 seconds, 16:9 landscape, 1080p.
Cinematic B-Roll, Urban Architecture
Variant: Veo 3.1 Fast · 8s · 16:9 landscape
Looking straight up at a glass-and-steel office tower from street level at midday. The camera slowly tilts from a steep upward angle to near-horizontal over the 8 seconds, catching cloud movement reflected in the building's surface. Harsh midday sunlight, 6000K. No human subjects. City ambient, distant traffic, light wind, no dialogue. 8 seconds, 16:9 landscape.
1080p Hero Shot, Luxury Fashion
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A tan leather tote bag rests on a cream-colored linen surface in a bright, window-lit room. The camera performs a slow rack focus from a blurred foreground flower arrangement to sharp focus on the bag's brass zipper detail. Soft natural window light from the right, 4500K. No human subjects. No audio. 8 seconds, 16:9 landscape, 1080p.
TikTok / Reels, Product Reveal 9:16
Variant: Veo 3.1 Fast · 8s · 9:16 portrait
A hand reaches into frame from below holding a small glass perfume bottle with a spherical cap. The hand lifts the bottle to center frame and holds it steady. Camera is locked off at eye level, centered composition. Bright clean studio light, 5500K. Subtle light flare off the glass bottle as it comes to rest. No dialogue. Quiet ambient studio. 8 seconds, 9:16 portrait.
Ambient Audio Scene, Nature
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
Early morning mist rises off a still lake surrounded by pine trees. The camera is stationary, framing the lake in a wide establishing shot. Dawn light, 2800K warm horizon glow, transitioning to 4000K as mist lifts. No human subjects. Natural ambient audio, water lapping gently, birds, wind through pines, no music, no dialogue. 8 seconds, 16:9 landscape, 1080p.
Ambient Audio Scene, Workspace
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A clean wooden desk with a half-open notebook, a pen, and a cup of black coffee. No person in frame. A ceiling fan turns slowly overhead. The camera holds a wide medium overhead shot looking down at 45 degrees. Warm diffused window light from the left, 3500K. Ambient audio, ceiling fan rotation, distant outdoor birdsong, no dialogue, no music. 8 seconds, 16:9 landscape, 1080p.
Animated Logo Reveal, Minimal
Variant: Veo 3.1 Fast · 8s · 16:9 landscape
A flat white logo mark, a simple geometric triangle, appears centered against a solid deep navy background. The logo fades in from 0 to 100% opacity over the first 3 seconds, then holds for 4 seconds, then fades out over the final second. No camera movement. No ambient audio. No human subjects. 8 seconds, 16:9 landscape.
Weather Plate, Overcast Sky
Variant: Veo 3.1 Standard · 8s · 16:9 landscape · 1080p
A slow-moving overcast grey sky, shot looking straight up with a slight wide-angle lens. Clouds drift slowly from left to right. No horizon line, no buildings, no objects, pure sky fill. Diffuse flat grey light. No subjects, no audio, no motion other than cloud drift. 8 seconds, 16:9 landscape, 1080p.
Veo 3.1 vs Other muvi.video Models
Related Links
The decision to reach for Veo 3.1 instead of another catalog model is driven by three variables: target output style, clip duration needs, and aspect-ratio range.
Veo 3.1 wins on commercial-grade photorealism and native ambient audio. Its 1080p output (1920×1080 landscape or 1080×1920 portrait) produces clean, detailed product hero shots and cinematic plates in 16:9 and 9:16. See the pricing page for current Veo 3.1 coin costs per clip. Note that Veo 3.1 does not support deterministic seed control: each generation is non-deterministic, so iteration runs through prompt refinement, not seed reuse.
Reach for a different model in three cases. First, when your script requires a single unbroken take longer than 8 seconds: Seedance 2.0 supports any integer duration from 4 to 15 seconds (lower resolution ceiling, 720p, or 1440p on the Start/End Frame variant), so it is the route when continuous duration matters more than pixel-level fidelity. Second, when you need an aspect ratio beyond 16:9 / 9:16, Seedance 2.0 supports 1:1, 4:3, 3:4, and 21:9 in addition to the standard pair. Third, when your reference aesthetic is tied to a specific stylized motion texture rather than commercial photorealism: Kling 3.0 (Text-to-Video, Image-to-Video, and Kling O3 variants) renders certain motion-smoothness aesthetics differently, test a clip before committing to a full sequence.
Veo 3.1 model details are at /models/veo-3.
Six Mistakes That Produce Weak Veo 3.1 Outputs
These are Veo 3.1-specific failure patterns, not generic AI video tips.
1. Over-prompting dialogue when ambient audio is the strength Veo 3.1's native audio strength is ambient matching, it reads your scene and generates contextually appropriate sound without instruction (a product claim per Google's Veo documentation). Writing lengthy dialogue instructions expecting lip-synced speech produces a scene where audio exists but mouths do not match. Veo 3.1's audio works best as ambient and environmental, rain, café murmur, city traffic, studio quiet. Specify it with one concise audio cue and let the model handle the rest. For shots that depend on synced spoken dialogue on screen, plan your audio in post rather than expecting the model to match lip movement.
2. Asking for stylized film grain on commercial-clean output Veo 3.1's default render is commercial-clean: sharp edges, accurate color, low noise. Adding "35mm film grain" or "VHS texture" to a product prompt fights the model's rendering bias and often produces inconsistent noise rather than styled grain. For commercial product shots, lean into Veo 3.1's natural output style. If you need a stylized look, validate on Veo 3.1 Fast first, the lower-fidelity variant is more tolerant of style requests.
3. Treating Veo 3.1 as a seed-deterministic model Veo 3.1 does not support deterministic seeds, each generation is non-deterministic. The most effective iteration pattern is to change exactly one phrase in your prompt per generation, then evaluate the delta. This keeps causality clean: if the output improves, you know which phrase drove it. For tighter frame-level consistency across takes, use the Veo 3.1 ECL Image-to-Video variant with a reference still from a prior generation as your composition anchor.
4. Combining two camera moves in one prompt Veo 3.1 executes one camera instruction accurately per generation. "Push in while tilting up" produces motion that splits the difference rather than performing both moves cleanly. Write one move. If your storyboard requires two angles, generate them as separate clips and cut in post.
5. Writing action instructions for the background Veo 3.1's attention on detailed background-motion instructions is weaker than its attention on foreground-subject instructions. "Background extras walk in both directions" in a street scene prompt often produces a static or loosely populated background. Write background context as environment description ("a busy street with pedestrians") rather than as an action instruction.
6. Setting the clip format last-minute The 8-second maximum means every prompt you write for Veo 3.1 should account for clip length at the design stage. A product rotation that requires 15 seconds of content needs to be storyboarded as two 8-second clips, not squeezed into one. If you realize this at prompt time, the brief has to change. Build 8-second thinking into your production planning before you start generating.
How to Iterate Veo 3.1 Clips Without Seed Control
Veo 3.1 does not support deterministic seeds, each generation is non-deterministic. Iteration therefore works through prompt refinement, not seed reuse. Here is the correct loop.
Write from the anatomy, then generate
On the first generation, write your full prompt using the seven-component anatomy in Section 3. Generate and evaluate the output completely before making any change.
Evaluate the output against one specific criterion
Watch the full clip before generating again. Identify exactly one thing to change: "the lighting is too flat, I want the shadow side darker" or "the product is too small in frame, I need a tighter push-in." Do not stack multiple changes, isolate one variable per generation cycle.
If the composition is close, change one phrase and regenerate
Copy the same prompt, change only the phrase that addresses your identified issue, and regenerate. You will get a related but non-identical output. Because Veo 3.1 is non-deterministic, the new generation may drift slightly in other areas, this is expected. Evaluate the new output against your one criterion, not against everything simultaneously.
If the composition is wrong, rewrite the brief
If the generation missed the brief entirely, wrong framing, wrong environment, wrong action, do not iterate on a broken prompt. Write a cleaner prompt from Step 1 and start a new thread.
Lock the winning prompt as the template
When a Veo 3.1 output clears your review bar, save the full prompt text as a named template. On the next campaign requiring a similar shot, start from that template and adjust only the product or scene variables. See the pricing page for current Veo 3.1 coin costs so you can budget iteration count against the brief.
Note: if your workflow requires holding a specific composition across multiple re-renders with precision, use the Veo 3.1 ECL Image-to-Video variant, feed a reference still as your spatial anchor rather than relying on prompt-only iteration.
Veo 3.1 Prompt Guide, Frequently Asked Questions
What resolution does Veo 3.1 support on muvi.video?+
Veo 3.1 generates at 1080p on muvi.video, 1920×1080 in landscape (16:9) and 1080×1920 in portrait (9:16). Clips run at 4, 6, or 8 seconds per generation. If your deliverable requires a longer unbroken take, consider Seedance 2.0 which supports any integer duration from 4 to 15 seconds (lower resolution ceiling, 720p, or 1440p on the Start/End Frame variant). For wider aspect-ratio options (1:1, 4:3, 3:4, 21:9), Seedance 2.0 is also the route.
Does Veo 3.1 have seed control on muvi.video?+
No. Veo 3.1 does not support deterministic seeds, each generation is non-deterministic. Iteration works through prompt refinement: change one phrase per cycle and evaluate the delta. For tighter frame-level consistency across takes, use the Veo 3.1 ECL Image-to-Video variant with a reference still as the spatial anchor.
How is Veo 3.1 billed on muvi.video?+
Veo 3.1 generations draw from your coin balance per clip. Coin costs vary by variant (Fast, Quality, ECL Image-to-Video, ECL Start/End). See the pricing page for current Veo 3.1 coin costs and what each subscription tier includes, pick the plan that matches your expected generation volume.
Does Veo 3.1 generate dialogue with lip sync?+
No. Veo 3.1 generates native ambient audio that matches scene context automatically, café ambient in a coffee scene, rain in an outdoor wet scene, studio silence when no audio cues are present (a product claim per Google's Veo documentation). It does not perform dialogue-to-lip-movement matching. If your brief requires a character delivering a specific spoken line, plan that audio in post rather than expecting the model to match lip movement. Veo 3.1's audio strength is ambient and environmental, not spoken dialogue.
What is Veo 3.1 best at compared to other models on muvi.video?+
Veo 3.1 leads the muvi.video catalog on three tasks: product and commercial photorealism, native ambient audio scene generation, and 1080p hero shot quality. Google DeepMind developed Veo 3.1 and muvi.video surfaces four variants, Fast, Quality, ECL Image-to-Video, and ECL Start/End, letting you trade render quality for generation speed or feed reference imagery for spatial control. For continuous clips longer than 8 seconds, Seedance 2.0 supports up to 15 seconds. For stylized motion textures outside Veo 3.1's commercial aesthetic, Kling 3.0 is the option.
When should I use Veo 3.1 Fast instead of Veo 3.1 Quality?+
Use Veo 3.1 Fast during the early creative exploration stage, when you are validating prompt direction, testing composition ideas, or generating weather plates and background elements where maximum fidelity is not required. Use Veo 3.1 Quality for final deliverable generations, product hero shots, and any clip where 1080p output quality and maximum commercial fidelity matter. The Quality variant produces higher output fidelity; Fast trades some of that fidelity for speed.
How should I handle the 8-second clip limit when my storyboard needs longer clips?+
Plan Veo 3.1 clips as 8-second building blocks at the storyboard stage, not at prompt time. For a 24-second brand film sequence, write three distinct 8-second prompts, each with its own subject state, camera position, and action instruction, then cut them together in your editing tool. Where continuity matters between clips, use consistent subject descriptions across adjacent prompts to hold visual consistency. If you need a single unbroken take longer than 8 seconds, use Seedance 2.0 (4–15 seconds, lower resolution ceiling) instead.
Start Generating with Veo 3.1 on muvi.video
Pick a prompt from Section 4, select a Veo 3.1 variant (Fast, Quality, or one of the ECL Image-to-Video options) in muvi.video Studio, and generate your first 1080p clip. No software to install, no queue to join.
No credit card required · Works in your browser · See pricing for Veo 3.1 subscription options
Related Pages
More Resources
Veo 3.1 Prompt Guide
Copy-ready Veo 3.1 prompts for muvi.video. Covers prompt anatomy, 10 real examples by use case, when to reach for Seedance or Kling, and the iteration loop.