Suggested searches

GPT Image Introduction GPT Image Free Guides GPT Image Developer Docs Midjourney Introduction Midjourney Free Guides Midjourney Developer Docs Google Nano Banana Introduction Google Nano Banana Free Guides Google Nano Banana Developer Docs Adobe Firefly Image Introduction Adobe Firefly Image Free Guides Adobe Firefly Image Developer Docs FLUX Introduction FLUX Free Guides FLUX Developer Docs Ideogram Introduction Ideogram Free Guides Ideogram Developer Docs Recraft Introduction Recraft Free Guides Recraft Developer Docs Stable Diffusion Introduction ByteDance Seedream Introduction Grok Imagine Image Introduction Google Veo Introduction Google Veo Free Guides Google Veo Developer Docs Runway Introduction Kling AI Introduction ByteDance Seedance Introduction ByteDance Seedance Free Guides ByteDance Seedance Developer Docs Luma AI Introduction Adobe Firefly Video Introduction Adobe Firefly Video Free Guides Adobe Firefly Video Developer Docs Hailuo AI Introduction PixVerse Introduction PixVerse Free Guides PixVerse Developer Docs Pika Introduction Pika Free Guides Pika Developer Docs Alibaba Wan Introduction LTX Video Introduction Grok Imagine Video Introduction Google Gemini Introduction Google Gemini Free Guides Google Gemini Developer Docs OpenAI GPT Introduction Claude Fable Introduction Claude Fable Free Guides Claude Fable Developer Docs DeepSeek Introduction Qwen Introduction Llama Introduction Codex Introduction Codex Free Guides Codex Developer Docs Cursor Introduction Cursor Free Guides Cursor Developer Docs OpenClaw Introduction OpenClaw Free Guides OpenClaw Developer Docs Perplexity Introduction ElevenLabs Introduction Suno Introduction Manus Introduction Claude vs ChatGPT: Compare Them Without Borrowed Numbers Claude vs Gemini: Comparing Two Assistants Honestly Is Adobe Firefly Free? The Free Membership, the Credits and the First Year Adobe Firefly Pricing: Plans, Generative Credits and the Cost of One Image Adobe Firefly Video Cost: Credits per Second and per Clip Is Adobe Firefly Video Free? Where the Free Plan Stops FLUX.2 Prompt Guide: The Structure Black Forest Labs Documents FLUX.2 vs Nano Banana Pro: What Each Vendor Actually Publishes Ideogram 4.0 Prompt Guide: The Documented Structure and How Text Renders Ideogram Pricing: Plans, Credits and the Free Limits Midjourney Prompt Guide: Structure, Elements and Edit Instructions Midjourney Pricing: Four Plans, GPU Hours and What a Job Costs Nano Banana Pro Pricing: What Google Publishes and What It Does Not Nano Banana Pro vs Midjourney: Which One Fits Your Workflow Pika Credits: What One Pika 2.5 Video Costs Is Pika Free? What the $0 Plan Actually Includes What PixVerse Credits Cost, and Which Ledger You Are Paying Is PixVerse Free? Two Free Tiers, and What Each One Costs You Recraft V4.1: The Eight Variants and Which to Pick Recraft Pricing and the Free Plan: What the Vendor Publishes Seedance Official Website: Which Entrances Are First-Party Seedance 2.5 vs 2.0: The Parameters ByteDance Publishes Veo 3.1 vs Sora 2: What Each Vendor Still Confirms Veo 3.1 vs Kling 3.0: What Each Vendor Publishes Is Claude AI Free? What the $0 Plan Actually Gives You Claude Pricing Plans: Pro vs Max 5x vs Max 20x How to Use Adobe Firefly (Adobe Firefly Image 5) Adobe Firefly API: Credentials, Endpoints and the Image5 Schema How to Use Adobe Firefly Video Adobe Firefly Video API: /v3/videos/generate Explained Designing UI Mockups and App Screens with GPT Image 2.5 Sticker Packs and Transparent Emoji with GPT Image 2.5 Nano Banana Pro Prompt Guide: The Official Frameworks Is Nano Banana Pro Free? What Google Actually Confirms How to Use Pika 2.5 Pika API: the official developer site, billing and endpoints How to Use PixVerse PixVerse API: Platform, Endpoints and Credits How to Use ByteDance Seedance 2.5 Seedance API: The Real Endpoint, Model ID and Fields Veo 3.1 Prompt Guide: The Seven Elements Google Names Veo 3.1 Price and Free Access: The Official Numbers How to Use Claude AI Claude API: Getting Started How to Use FLUX AI: A FLUX.2 Getting-Started Guide FLUX.2 API: Getting Started with Black Forest Labs GPT Image 2.5 vs DALL·E 3 Making Infographics and Diagrams with GPT Image 2.5 How to Use Ideogram Ideogram API: Access, Endpoints and Text Rendering How to Use Midjourney V8.2 Midjourney API: What Exists and What Does Not How to Use Nano Banana Pro Nano Banana Pro API: Getting Started How to Use Recraft Recraft API: Access, Endpoints and Style Consistency How to Use Veo 3.1 Veo 3.1 API Pricing and Vertex AI Access GPT Image 2.5 Prompt Sharing GPT Image 2.5 vs Nano Banana 2 GPT Image 2.5 vs Seedream 5.0 Pro GPT Image 2.5 vs FLUX 2 GPT Image 2.5 vs Ideogram Is GPT Image 2.5 Free GPT Image 2.5 Sketch GPT Image 2.5 Templates GPT Image 2.5 Comment Editing GPT Image 2.5 Character Consistency GPT Image 2.5 Combine Images GPT Image 2.5 Text Rendering What Is GPT Image 2.5 GPT Image 2.5 API Overview Codex vs. ChatGPT: Which Should You Use? Codex app, CLI, IDE, or cloud: how to choose the right surface Your First Low-Risk Coding Task with Codex: A Safe Walkthrough What Is Cursor? Its AI Coding Workflow Explained Cursor vs. VS Code: Which Editor Fits Your Workflow? Cursor Features Explained: Agent, Tab, Context, and More What Is Gemini? Apps, Models, AI Studio, and API Explained Gemini Apps vs. Gemini API: How to Choose the Right Tool for the Job What Can Gemini Do? A Practical Capability Guide OpenClaw Foundation Explained: Governance and Independence OpenClaw Skill Workshop Guide: Review Reusable Workflows OpenClaw Skill Cards: Read ClawHub Security Scans OpenClaw 2.0 Guide: New Features and Upgrade Checks OpenClaw LTS Guide: Choosing extended-stable or stable Install OpenClaw: Desktop, Script, npm, and Source Options OpenClaw Node.js Setup: Versions, Installation, and PATH How to Write Better Codex Prompts: A Practical Framework How to Review Codex Code Changes Before You Commit Cursor Beginner Tutorial: From Install to First Reviewed Edit Cursor Rules Tutorial: Project Rules, User Rules, and AGENTS.md Install Cursor on Windows and Configure a Chinese Interface Cursor MCP Tutorial: Configure, Verify, and Secure MCP Servers Gemini Prompt Guide: Better Instructions and Templates Gemini API Quickstart: Key, Python SDK and First Call Gemini Web App Guide: Login, Files, Chats and Privacy Gemini API Key Security: Storage, Restrictions and Rotation What Is Codex? Capabilities, Limits, and Ways to Use It Codex Beginner Tutorial: Complete Your First Safe Task Install Codex CLI: Sign In and Run Your First Safe Task Codex AGENTS.md Guide: Layered Rules and Validation Codex CLI Commands: Sessions, Review, and Automation How to Use Gemini: Web, Android & iPhone Setup Gemini Features Guide: Chat, Files, Images & Live How to Chat with Gemini: Prompts, Follow-Ups & Live Gemini AI Image Generator Guide: Prompts & Editing Gemini vs GPT-4: Features, Limits & Which to Use Gemini AI Assistant Guide: Mobile, Apps & Privacy Gemini Prompt Engineering Guide: Patterns & Examples Gemini Chat API Guide: Multi-Turn Prompts in Python Gemini System Instructions: API Guide & Examples Gemini Context Caching Guide: Cost, Latency & API
AI Tool Blog Veo 3.1 learning hub

Generate video with Veo 3.1.

Google Veo 3.1 is DeepMind’s current video generation model. It turns a written prompt, a single image or a set of reference images into a short clip, and generates the dialogue, sound effects and ambient audio in the same pass as the picture.

Every capability, benchmark summary and access route on this page comes from Google DeepMind’s own Veo model page, and the showcase stills are the demo frames Google publishes on that page.

Official examples

What Google’s own Veo demos render

These stills are the frames Google publishes beside its own demo clips in the model page’s showcase section, and each prompt below is the one Google wrote for that clip. They are the vendor’s own demonstrations rather than anything generated for this site.

Google Veo demo still of an owl flying over a moonlit forest
A follow shot of a wise old owl over a moonlit forest while a nervous badger speaks. Google’s prompt hands the badger its line, and the clip is one of the page’s dialogue examples — script and picture from one prompt.
Google Veo demo still of a detective questioning a rubber duck on a desk
A detective interrogating a nervous-looking rubber duck. The prompt Google wrote supplies the detective’s line, which is the point of the example: the spoken part is generated with the picture, not added afterwards.
Google Veo demo still of two women walking on an Irish coastal headland
Two women walking an Irish headland in the 1860s, “their long, modest dresses of homespun fabric whipping gently in the strong coastal wind”. Google’s own prompt, and its period-drama register.
Google Veo demo still of a wax figure holding a candle
A small pale-yellow figure crafted from wax, holding a candle in a warm interior. The prompt names both the material and the light, which is the pair this frame has to get right.
Google Veo demo still of a keyboard with keys made of candy
A keyboard whose keys are made of different kinds of candy. Google’s prompt asks for “sweet, crunchy sounds” as the typing audio, so the clip is about sound design as much as it is about the object.
Google Veo demo still of crystalline flowers on an iridescent plain
A snow-covered plain of iridescent moon-dust under twilight skies, with thirty-foot crystalline flowers blooming above it — one of the more surreal clips in the showcase, and the widest in scale.

The demo stills below belong to Google and are reproduced with credit, from the Veo model page. Google DeepMind’s Veo model page

What it is

Veo 3.1 in Google’s own words

Google introduces Veo 3.1 as its leading video generation model, “designed to empower filmmakers and storytellers”, and builds the page around a single claim: that video and audio now arrive together. The showcase, the control sections and the benchmark summaries below all come from that page.

01

Video and audio in one pass

Dialogue, sound effects and ambient noise are generated natively rather than dubbed on afterwards. The page’s own line for it is “video, meet audio”, and several of its demo prompts write the spoken part into the prompt itself.

02

Control at the level of the shot

Past a text prompt, the model accepts reference images for a scene, a character or an object, a style reference, a first and a last frame, and camera instructions such as zoom in or move right.

03

Edits rather than restarts

Google documents outpainting an existing frame, adding an object to footage and removing one from it, extending a clip from its final second, and driving a character’s performance with your own body, face and voice.

04

Resolutions aimed at an edit

Output is documented at 1080p and 4K. Google presents 1080p as the resolution for material that has to cut against other footage, and 4K as the one for texture and detail.

What it does

The controls Google documents

Google groups the model’s control surface under headings of its own and demonstrates each one with a clip. Read together they answer the practical question — what can I actually steer — and every item below is a heading on the model page rather than an interpretation of one.

01

Ingredients to video

Supply reference images of a scene, a character or an object and the shot is built around them. The page’s own prompts for this section ask for a commercial, a film trailer and a music video from the same set of references.

02

Match your style

A style reference carries the look, so an aesthetic that is hard to put into words — a paper diorama, a painting, a particular grade — can be handed over as an image instead of described.

03

Keep your characters consistent

Character references hold a subject’s appearance steady across separate shots, which is the piece that makes a multi-shot sequence possible rather than a set of unrelated clips.

04

Extend your scene

A clip can be continued from its last second while keeping the look and the sound consistent; the page shows prompts chained this way into one longer piece.

05

Camera controls

Framing and movement are set explicitly rather than asked for in prose. The page names four: move back, zoom in, move up and move right.

06

First and last frame

Two images are enough to define a transition, with the model filling in the motion between them.

07

Outpainting, adding and removing objects

Outpainting extends the frame outward so one clip can fit a different screen shape, while separate controls add an object to a scene or take one out, accounting for scale, interaction and shadow.

08

Character and motion controls

A performance can be driven by your own body, face and voice, and an object’s movement can be defined by drawing the path you want it to take.

09

1080p and 4K output

Google documents two output resolutions and says which is for what: 1080p for a sharper, cleaner file to edit with, 4K where texture and detail matter more than weight.

Documented specifications

What Google publishes about Veo 3.1

Only facts Google states on its own pages appear here. Where the model page gives no figure — a maximum clip length, a per-second price — this table does not invent one.

Developer
Google DeepMind
Current model
Veo 3.1
Category
Video generation
Output
Video with natively generated audio
Output resolutions
1080p and 4K
Documented inputs
Text prompts, images, reference images, video
Documented controls
Ingredients, style reference, character consistency, scene extension, camera controls, first and last frame, outpainting, object add and remove, character and motion controls
Watermarking
SynthID
Consumer surfaces
Gemini, Google Flow, Google Vids
Developer surfaces
Google AI Studio, Gemini API
Benchmarks last updated
October 2025
Stated limitation
Google says natural and consistent spoken audio, particularly for shorter speech segments, remains an area of active development.
How to use it

Documented access routes

Veo 3.1 is a hosted model, so every route is an account on someone else’s service. Google names five surfaces on the model page, and each row below links to the one Google itself points at.

Gemini

The consumer assistant, where Veo appears as a named video surface.

Open

Google Flow

Google’s AI filmmaking tool, built for assembling clips into scenes and stories.

Open

Google Vids

Video creation inside Google’s work suite, aimed at workplace output rather than film.

Open

Google AI Studio

The browser path from prompt to API key, with a Veo studio app of its own.

Open

Gemini API

The developer route, documented on ai.google.dev with worked prompt examples.

Open
FAQ

Veo 3.1 questions people search for

What is Veo 3.1?

Google DeepMind’s current video generation model. It produces short clips with dialogue, sound effects and ambient audio generated in the same pass as the picture, from a text prompt, a single image or a set of reference images.

Does Veo generate audio as well as video?

Yes, and it is the model’s headline claim. Google’s own wording is “video, meet audio”, the showcase section is built around it, and several of the demo prompts write the spoken line into the prompt rather than leaving it to be dubbed later.

What resolution does Veo 3.1 output?

Google documents 1080p and 4K. Its page presents 1080p as the sharper, cleaner option for editing and 4K as the one for texture and detail.

Can Veo 3.1 edit a video I already have?

It can work on footage you supply. The documented controls include outpainting to extend the frame, adding an object, removing an object, extending a clip from its last second, and setting the first and last frame of a transition.

How long are Veo 3.1 clips?

The model page does not state a single maximum duration. It documents scene extension as the way to build longer pieces, and the footnotes to its head-to-head comparisons describe the evaluation clips as 6 or 8 seconds long.

Is Veo output watermarked?

Yes. Google says videos made with Veo are marked with SynthID, its watermarking and detection technology for AI-generated content, and that outputs also go through safety evaluations and memorisation checks before release.

Where can I use Veo 3.1?

Google names five surfaces: Gemini, Google Flow, Google Vids, Google AI Studio and the Gemini API. There is no self-hosted or open-weights route.

Published library

Latest Google Veo 3.1 articles

Browse every published Google Veo 3.1 article, from introductions to practical guides and developer documentation.