AI Potluck
Back to Gap Map Model components / Image, video, 3D & music generation

Veo

Google
closed / Overall score: n/a

Veo is Google DeepMind's video generation model line, which generates video with native audio from text and image prompts. It is served through the Gemini API and Google's Flow filmmaking tool in standard, Fast and Lite tiers, with 4K output and portrait formats.

Gemini Omni Flash, which Google presents as a separate Gemini model family and its default for video in the Gemini API, is not part of this entry.

Openness

1 high confidence
1.0
weights
closed(no weights are distributed
data
closed
code
closed
license
Proprietary(Gemini API terms, which bar attempts to extract or replicate the model)

Veo is available only through Google's API and apps, and Google's terms forbid extracting the model. No weights, training data or code are published.

Adoption

not assessed

Google publishes a count of videos made in its Flow tool but no usage figure for Veo itself, and a hosted model has no download count, so no reading was possible.

Capability

3 medium confidence
3.0

Veo 3.1 places in the upper half of the video arena, level with LTX-2.5 on the silent board. Google's newer Gemini Omni Flash and the downloadable MiniMax H3 both place well above it.

Verified 2026-09-26