How to choose
The decision is usually not "which model is best" but "which model is right for this
shot". Three questions settle almost every case.
Does a recurring person appear in it? If yes, use Kling o3. It accepts
the locked-cast composite, which is the only mechanism here that binds identity
structurally. Seedance 2.0 will refuse the shot outright, and Veo 3.1 and Grok will
render a plausible stranger.
Is it a hero shot? Veo 3.1 is the highest-fidelity option and brings
native audio, at 45 credits per second and durations snapped to 4, 6, or 8 seconds. On a
six-second hero shot that is 270 credits — worth it once per film, rarely worth it
twelve times.
Are you still exploring? Grok Imagine 1.5 at 20 credits per second is
the cheapest way to see a first frame move, as long as you accept 720p and the no-refund
caveat. Promote the shot to a better model once the composition is settled.
What a real project costs
A 30-second ad with eight shots, averaging six seconds each, is 48 seconds of finished
video. On Kling o3 at 20 credits per second that is around 960 credits. The same 48
seconds on Veo 3.1 is 2,160. Mixing them — Veo on the two hero shots, o3 on the other
six — lands near 1,300, which is the practical reason per-shot model choice exists.
Add lip-sync to the four shots with dialogue and you add 15 credits per second across
roughly 24 seconds: 360 credits. Every plan on the
pricing page lists its monthly credit inclusion, and the project
math there works through a 90-second film and a localised variant.
Why the model is not the product
Every model on this page is available to everyone. A subscription to any one of them
gets you a prompt box and a clip. What it does not get you is the thing that makes eight
clips cut together: a locked cast, an approved look, locations that persist, and a
script the shot list actually reads from.
That is the part GATA is. Models are interchangeable renderers inside it — see
script to video, in parallel for the workspace
the renderers plug into, and
shot inheritance for the mechanism that lets
you swap a shot's model without redoing anything around it.
Frequently asked
Which AI video model is best for character consistency?
Kling o3 Pro. It is the only pair of models — o3 and v3 — that accept locked-cast composites as references, so the approved face is bound to the generation rather than described in a prompt. It is also GATA's default model for that reason. Veo 3.1, Seedance 2.0, and Grok Imagine 1.5 do not take character references.
How much does a second of AI video cost in GATA?
Between 20 and 45 credits per second depending on the model: Kling o3 and Grok Imagine 1.5 are 20, Kling v3 is 25, and Seedance 2.0 and Veo 3.1 are 45. Optional uplifts add per second on top — 4K adds 30, lip-sync adds 15, background noise adds 5.
Can I use a different model for each shot?
Yes. The model is a per-shot choice, not a per-project one. Because script, cast, look, and locations live in project state, switching a shot's model does not re-do any of them — the shot inherits the same context and only the renderer changes.
Why does Seedance 2.0 refuse shots with people in them?
Seedance 2.0 rejects generations that reference faces. That is a model-side restriction, not a GATA setting. If your script has human characters, GATA warns you before the project starts and you should pick Kling o3 instead.
Which models can render 4K?
Kling o3, Kling v3, and Veo 3.1. The 4K toggle adds 30 credits per second and is hidden for Seedance 2.0 and Grok Imagine 1.5, which cannot render it. Finished scene clips can also be upscaled at the master stage for 10 credits per second.
What happens to my credits if a generation fails?
Failed generations are refunded, with one disclosed exception: xAI charges for every Grok Imagine 1.5 request including ones it blocks for policy reasons, so a blocked Grok clip is not refunded. Every other model refunds a content-policy failure.