Veo 3.1 AI
Google DeepMind's True-4K Video Model with Native Audio & Lip Sync
Veo 3.1 is the first mainstream AI video model to deliver true 4K output, with native audio, tighter lip sync than Veo 3, reference-to-video, and extend up to ~148 seconds. Start with the Fast tier at 54 credits, or unlock the audio tier at 80 credits.
Veo 3.1 Video Showcase
True 4K AI video from Google DeepMind

Introducing Veo 3.1 from Google DeepMind
Launched in January 2026, Veo 3.1 is Google DeepMind's flagship video model and the first mainstream AI video generator with true 4K output. It improves on Veo 3 with better audio/video sync, native audio with lip sync, reference-to-video control, and an extend feature that chains clips up to roughly 148 seconds. Choose the Fast tier at 54 credits or the audio tier at 80 credits.
- First Mainstream True 4KGenuine 4K resolution - a first for mainstream AI video generation
- Reference-to-VideoSupply reference images to control characters, objects, and style
- Extend to ~148 SecondsChain generations into long, coherent sequences with extend
Veo 3.1 Features
Why Veo 3.1 sets the new standard for AI video
True 4K Output
The first mainstream AI video model to deliver genuine 4K resolution
Native Audio + Lip Sync
Dialogue, ambience, and effects generated with accurate lip sync
Better A/V Sync than Veo 3
Tighter alignment between sound and picture in every generation
Reference-to-Video
Guide generations with reference images for consistent characters
Extend Up to ~148s
Build multi-minute scenes by extending clips seamlessly
Fast & Audio Tiers
Fast tier at 54 credits, audio tier at 80 credits - pick per project
Veo 3.1 Frequently Asked Questions
Everything you need to know about Google DeepMind's Veo 3.1
What is Veo 3.1?
Veo 3.1 is Google DeepMind's flagship AI video model, released in January 2026. It's the first mainstream AI video generator with true 4K output, featuring native audio with lip sync, reference-to-video, and an extend feature that reaches up to roughly 148 seconds.
How much does Veo 3.1 cost?
Veo 3.1 Fast costs 54 credits per 8-second generation. The audio tier, which adds native audio generation with lip sync, costs 80 credits.
Veo 3.1 vs Veo 3: what's improved?
Veo 3.1 upgrades Veo 3 in four key ways: true 4K resolution (a mainstream first), noticeably better audio/video synchronization, reference-to-video support for character and style control, and the extend feature that chains clips up to ~148 seconds. Lip sync accuracy is also significantly improved over Veo 3.
Does Veo 3.1 really generate 4K video?
Yes. Veo 3.1 is the first mainstream AI video model to offer genuine 4K output, delivering detail that holds up on large displays and in professional editing pipelines.
What is reference-to-video in Veo 3.1?
Reference-to-video lets you supply one or more reference images - a character, product, or location - and Veo 3.1 keeps them consistent in the generated video. Combined with extend, it's ideal for longer narrative projects.
How long can Veo 3.1 videos be?
A single Veo 3.1 generation is 8 seconds. Using the extend feature, you can chain generations into continuous sequences of up to approximately 148 seconds - over two minutes of coherent AI video.
Experience Veo 3.1 in True 4K
Google DeepMind's flagship video model - Fast tier from 54 credits



