Prompt Workbench
Free Veo Prompt Builder
Start from an ad brief, not a blank prompt box. Build a short shot plan, a paste-ready Veo prompt, and a structured JSON version in your browser.
Premium
AI Brief Assist
Paste a rough idea, product note, or messy offer and let AI fill the brief fields before you tune the final prompt.
How the Veo prompt builder works
Veo does not fail because your idea is bad. It fails because a sentence like “a woman drinking coffee, cinematic” leaves the model to invent the shot, the lens, the light, the pacing, and the sound. Every field it invents is a field that changes on the next generation. This builder’s job is to close those gaps before you spend a credit — it takes a short brief and assembles the six parts a video model actually reads: subject, action, setting, camera, lighting, and audio.
It runs entirely in your browser. Nothing is uploaded, there is no account, and there is no generation cost, because the tool writes prompts rather than video. You take the output to Veo, or to whichever model you run.
Text output vs JSON output
The builder produces the same prompt in two shapes, and the right one depends on how you work.
- Text is the paste-ready prompt for the Veo app, Gemini, or any chat-style interface. One prose block, ordered so the most important elements come first — models weight early tokens more heavily. Use it for one-off shots.
- JSON is the same prompt split into named fields. Use it
when you are producing a batch and need shots to match: you can change
subjectacross ten variants whilecamera,lighting, andnegativestay byte-identical. It also diffs cleanly in version control, which matters once a prompt becomes an asset you maintain rather than a message you send.
Neither is “better”. JSON is not a secret mode that unlocks quality — it is the same instruction with structure you can hold steady. The gain is repeatability, not fidelity.
A worked example
What most people write:
A person using our skincare product. UGC style, vertical, 8 seconds. What the builder assembles from the same brief:
Handheld selfie video, 9:16 vertical, 8 seconds. A woman in her late
twenties sits on the edge of a bed in a sunlit bedroom, holding a
frosted-glass serum bottle up to camera. She glances at the bottle,
then back to lens, and says: "Three weeks. That's it." Natural
window light from frame left, soft shadows, slight lens breathing
from the handheld grip. Room tone and a faint street hum underneath
her voice. No on-screen captions, no background music, no zoom. The second version is not longer for its own sake. Every clause removes a decision the model would otherwise make for you: the aspect ratio fixes the crop, the named light direction fixes the mood across takes, the spoken line fixes lip-sync timing, and the closing negatives kill the three defaults Veo most often adds unasked — captions, stock music, and a drifting zoom.
Which models this works with
Prompts are built for the current Veo line — Veo 3.1, Veo 3.1 Fast, and Veo 3.1 Lite. Veo 4 has not shipped: at I/O 2026 Google announced Gemini Omni Flash as its next-generation video model rather than a Veo 4 release. Because the text output is plain prose rather than a proprietary syntax, it transfers to most other video models with light editing — the structure is what carries, not the formatting. Audio direction and negative prompts are the two parts most likely to need adjusting when you move between models.
Where to go next
If you would rather start from something already proven than from a blank brief, the Veo prompt library collects working structures by use case, each labelled with how it was verified. The closest starting points to this builder are the UGC-style prompts, product video prompts, and image-to-video prompts.
If you want the reasoning rather than the recipe, the guides cover the prompt formula and structure, camera movement language, the JSON prompt format, and negative prompts for troubleshooting.