LiveVeo 3.1 is live: reference images, 4K output and scene extension up to ~148 seconds. Try it now →

3
The model that gave AI video a voice

Veo 3
Google's AI video model.

Veo 3 turns text or an image into a finished video with the sound already in: score, effects and lip-synced dialogue generated in the same pass as the picture. Generate text-to-video and image-to-video in the browser.

Try Veo 3 freeAvailable in the studio alongside the new Veo 3.1.
Overview

What is Veo 3?

Veo 3 is an AI video generation model from Google DeepMind. It was the first mainstream model to generate video and synchronized audio in a single pass — put a line of dialogue in quotes and the character says it with matching lip movement. It reads plain-language direction for camera moves, lighting and tone, and renders physics that hold up to a rewatch.

Released in May 2025 at Google I/O, Veo 3 set the template every video model since has chased, and it remains available on Veo Studio alongside its successor, Veo 3.1. You get a simple prompt box, no downloads, and a free tier to start. Developers can also call it through the Veo API.

Features

Veo 3 features.

Prompting

Veo 3 speaks the language of film.

Camera moves, lighting setups, lens choices, emotional tone, write a Veo prompt the way a director would and the model executes it. No settings panels, just words.

PromptSlow dolly-in on a street racer leaning on her car at night, neon rim light, rain just ended. She smirks and says "you're already too late." Cut to wide as the engine roars, bass-heavy score kicks in.
Getting started

How to use Veo 3.

  1. 01

    Sign in and open the studio

    Create a free account and pick Veo 3 from the model selector (Veo 3.1 is the default), no waitlist, no invite code and nothing to download.

  2. 02

    Write a prompt or add an image

    Describe the shot, the camera move and any dialogue in quotes. Optionally attach a photo as the starting frame for image-to-video.

  3. 03

    Generate and download

    Veo 3 returns a finished clip of 4 to 8 seconds at up to 1080p, with sound. Pick 16:9 or 9:16 for TikTok, Reels, Shorts or ads and export.

Release date

Veo 3 release date.

Google DeepMind released Veo 3 in May 2025 at Google I/O, alongside a faster, cheaper Veo 3 Fast variant. It followed Veo 2 (late 2024), and its headline feature — native audio generated with the picture — made it the most talked-about model release of the year.

Its successor, Veo 3.1, launched in October 2025 with reference images, first-and-last-frame control, 4K output and scene extension. On Veo Studio it's now the default model for every plan, and Veo 3 remains available as an option.

Spec sheet

Veo 3 specifications.

Clip length
4–8 seconds
Resolution
720p · 1080p
Aspect ratios
16:9 · 9:16
Inputs
text · image (first frame)
Audio
native: music, SFX, dialogue
Provenance
SynthID invisible watermark

Veo 3.1 is here.

Veo 3.1 adds up to 3 reference images, first-and-last-frame control, 4K output and scene extension up to ~148 seconds, on every plan, at no extra cost. It's live now.

Explore Veo 3.1
FAQ

Veo 3 FAQ.

Start creating with Veo 3.

The free plan gets you a couple of videos a month, enough to see what Veo 3 does with your idea.

Make your first video