Skip to main content
The prompt is the whole interface to the Video Agent. The agent makes every decision you leave open, so the more direction you give, the closer the first cut lands to what you pictured. A good prompt reads like a creative brief: what the video is, how long it runs, and what it should look like. This page builds one real prompt step by step — and ends with the video it produced.

Put the duration in the prompt

The agent paces scenes to the length you ask for. Lead with it: “Make a 30-second portrait real estate market update…” produces a video paced to 30 seconds, scene by scene.

Add a style paragraph

Describe the video you want, then add a short style paragraph: what the graphics should look like, the colors, the energy, how things move. The agent reads it and makes every visual decision through it — anything you leave out, it decides for you.
Example prompt
A style paragraph that works has a consistent anatomy: a name for the style, the exact palette (hex codes and all), the art direction, how things move, what the transitions are, and one closing line for the vibe. Five or six sentences, and every graphic in the video has a rulebook — because scenes are built in code with Hyperframes, the agent follows it literally. Prefer a pre-built look instead? Pass a style_id — see Styles & References.

Paste your script for exact control

If you already have a script, paste the whole thing in and the agent follows it scene by scene, building the visuals around your exact words. Combine it with the style paragraph and a duration for the tightest brief:
Example prompt with script

The output

This is the video that prompt produced — one prompt, one pass:

Generated by Video Agent from the prompt above.

The first cut is a draft. To close the gap to post-ready, keep talking to the agent in chat mode — revise a single scene (“change the background of scene two”) while every other scene holds still. Prompt for the 90, edit for the 10.

More levers

  1. Attach reference files. Pass slides, images, or documents in the files array — the agent folds them into the video as visual context or content sources. See Upload Assets.
  2. Pin a specific avatar or voice with avatar_id (see Avatars) and voice_id (see Browse Voices) for consistency, or omit them to let the agent choose.
  3. Apply your brand once with brand_kit_id from GET /v3/brand-kits — colors, fonts, and logo ship on-brand by default on every video after.
  4. Set orientation when you know the target platform ("portrait" for mobile/social, "landscape" for presentations).
  5. Review the plan before it builds. With "mode": "chat", the agent comes back with a scene-by-scene plan you can push back on before production starts. See Interactive Sessions.
For more worked examples, see the Video Agent prompt guide on the blog.