Why “one shot” is the unit that matters
Text-to-video models are trained to render a continuous shot, and Vizify’s video capability returns exactly one clip per request. A prompt that describes a sequence (“she opens the door, walks to the window, then turns and smiles”) asks the model to compress three shots into a few seconds, and it will usually blur them together. Write the shot you need now, generate it, and write the next shot as its own prompt.
A reliable order for the brief is subject, action, setting, camera, light, style, audio. You do not need headings or a template, but covering each of those in plain sentences removes most of the guessing the model would otherwise do.
Model limits change what a good prompt looks like
Because every Vizify video model exposes its own parameter set, the same brief behaves differently depending on where it runs:
- Kling 3 is the current default. It runs 3 to 15 seconds, offers 720p, 1080p, or 4K, accepts up to two reference images, and keeps audio off unless you turn it on.
- Veo 3.1 is fixed at 8 seconds and 720p and always returns audio, so write the audio line with care and do not ask it for a silent clip.
- Seedance 2 is always silent but goes up to 4K and 15 seconds, which suits product and B-roll shots you will score later.
- Seedance 2.5 runs 4 to 30 seconds with optional audio and accepts many more reference images, which helps when a scene has several fixed elements.
- Grok Imagine Video 1.5 always includes audio and takes up to seven reference images.
- Hailuo 2.3 needs exactly one reference image, and it supports 1080p only on 6-second clips.
If a detail in your prompt conflicts with the model’s limits, such as a 20-second request to a model capped at 15 or “silent” on a model that always returns audio, the model’s contract wins. Check the model page before you write the timing and audio lines.
Reviewing and revising
Watch the clip at full size from start to finish. The first things to drift are faces, hands, product geometry, logos, and any small text. When a revision is needed, keep everything else in the prompt the same and change one line, whether that is the camera move, the light, or the action, so you learn what actually made the difference. Each revision is a new generation and uses credits the same way the first one did.