Suggested searches

GPT Image Introduction GPT Image Free Guides GPT Image Developer Docs Midjourney Introduction Midjourney Free Guides Midjourney Developer Docs Google Nano Banana Introduction Google Nano Banana Free Guides Google Nano Banana Developer Docs Adobe Firefly Image Introduction Adobe Firefly Image Free Guides Adobe Firefly Image Developer Docs FLUX Introduction FLUX Free Guides FLUX Developer Docs Ideogram Introduction Ideogram Free Guides Ideogram Developer Docs Recraft Introduction Recraft Free Guides Recraft Developer Docs Stable Diffusion Introduction ByteDance Seedream Introduction Grok Imagine Image Introduction Google Veo Introduction Google Veo Free Guides Google Veo Developer Docs Runway Introduction Kling AI Introduction ByteDance Seedance Introduction ByteDance Seedance Free Guides ByteDance Seedance Developer Docs Luma AI Introduction Adobe Firefly Video Introduction Adobe Firefly Video Free Guides Adobe Firefly Video Developer Docs Hailuo AI Introduction PixVerse Introduction PixVerse Free Guides PixVerse Developer Docs Pika Introduction Pika Free Guides Pika Developer Docs Alibaba Wan Introduction LTX Video Introduction Grok Imagine Video Introduction Google Gemini Introduction Google Gemini Free Guides Google Gemini Developer Docs OpenAI GPT Introduction Claude Fable Introduction Claude Fable Free Guides Claude Fable Developer Docs DeepSeek Introduction Qwen Introduction Llama Introduction Codex Introduction Codex Free Guides Codex Developer Docs Cursor Introduction Cursor Free Guides Cursor Developer Docs OpenClaw Introduction OpenClaw Free Guides OpenClaw Developer Docs Perplexity Introduction ElevenLabs Introduction Suno Introduction Manus Introduction Claude vs ChatGPT: Compare Them Without Borrowed Numbers Claude vs Gemini: Comparing Two Assistants Honestly Is Adobe Firefly Free? The Free Membership, the Credits and the First Year Adobe Firefly Pricing: Plans, Generative Credits and the Cost of One Image Adobe Firefly Video Cost: Credits per Second and per Clip Is Adobe Firefly Video Free? Where the Free Plan Stops FLUX.2 Prompt Guide: The Structure Black Forest Labs Documents FLUX.2 vs Nano Banana Pro: What Each Vendor Actually Publishes Ideogram 4.0 Prompt Guide: The Documented Structure and How Text Renders Ideogram Pricing: Plans, Credits and the Free Limits Midjourney Prompt Guide: Structure, Elements and Edit Instructions Midjourney Pricing: Four Plans, GPU Hours and What a Job Costs Nano Banana Pro Pricing: What Google Publishes and What It Does Not Nano Banana Pro vs Midjourney: Which One Fits Your Workflow Pika Credits: What One Pika 2.5 Video Costs Is Pika Free? What the $0 Plan Actually Includes What PixVerse Credits Cost, and Which Ledger You Are Paying Is PixVerse Free? Two Free Tiers, and What Each One Costs You Recraft V4.1: The Eight Variants and Which to Pick Recraft Pricing and the Free Plan: What the Vendor Publishes Seedance Official Website: Which Entrances Are First-Party Seedance 2.5 vs 2.0: The Parameters ByteDance Publishes Veo 3.1 vs Sora 2: What Each Vendor Still Confirms Veo 3.1 vs Kling 3.0: What Each Vendor Publishes Is Claude AI Free? What the $0 Plan Actually Gives You Claude Pricing Plans: Pro vs Max 5x vs Max 20x How to Use Adobe Firefly (Adobe Firefly Image 5) Adobe Firefly API: Credentials, Endpoints and the Image5 Schema How to Use Adobe Firefly Video Adobe Firefly Video API: /v3/videos/generate Explained Designing UI Mockups and App Screens with GPT Image 2.5 Sticker Packs and Transparent Emoji with GPT Image 2.5 Nano Banana Pro Prompt Guide: The Official Frameworks Is Nano Banana Pro Free? What Google Actually Confirms How to Use Pika 2.5 Pika API: the official developer site, billing and endpoints How to Use PixVerse PixVerse API: Platform, Endpoints and Credits How to Use ByteDance Seedance 2.5 Seedance API: The Real Endpoint, Model ID and Fields Veo 3.1 Prompt Guide: The Seven Elements Google Names Veo 3.1 Price and Free Access: The Official Numbers How to Use Claude AI Claude API: Getting Started How to Use FLUX AI: A FLUX.2 Getting-Started Guide FLUX.2 API: Getting Started with Black Forest Labs GPT Image 2.5 vs DALL·E 3 Making Infographics and Diagrams with GPT Image 2.5 How to Use Ideogram Ideogram API: Access, Endpoints and Text Rendering How to Use Midjourney V8.2 Midjourney API: What Exists and What Does Not How to Use Nano Banana Pro Nano Banana Pro API: Getting Started How to Use Recraft Recraft API: Access, Endpoints and Style Consistency How to Use Veo 3.1 Veo 3.1 API Pricing and Vertex AI Access GPT Image 2.5 Prompt Sharing GPT Image 2.5 vs Nano Banana 2 GPT Image 2.5 vs Seedream 5.0 Pro GPT Image 2.5 vs FLUX 2 GPT Image 2.5 vs Ideogram Is GPT Image 2.5 Free GPT Image 2.5 Sketch GPT Image 2.5 Templates GPT Image 2.5 Comment Editing GPT Image 2.5 Character Consistency GPT Image 2.5 Combine Images GPT Image 2.5 Text Rendering What Is GPT Image 2.5 GPT Image 2.5 API Overview Codex vs. ChatGPT: Which Should You Use? Codex app, CLI, IDE, or cloud: how to choose the right surface Your First Low-Risk Coding Task with Codex: A Safe Walkthrough What Is Cursor? Its AI Coding Workflow Explained Cursor vs. VS Code: Which Editor Fits Your Workflow? Cursor Features Explained: Agent, Tab, Context, and More What Is Gemini? Apps, Models, AI Studio, and API Explained Gemini Apps vs. Gemini API: How to Choose the Right Tool for the Job What Can Gemini Do? A Practical Capability Guide OpenClaw Foundation Explained: Governance and Independence OpenClaw Skill Workshop Guide: Review Reusable Workflows OpenClaw Skill Cards: Read ClawHub Security Scans OpenClaw 2.0 Guide: New Features and Upgrade Checks OpenClaw LTS Guide: Choosing extended-stable or stable Install OpenClaw: Desktop, Script, npm, and Source Options OpenClaw Node.js Setup: Versions, Installation, and PATH How to Write Better Codex Prompts: A Practical Framework How to Review Codex Code Changes Before You Commit Cursor Beginner Tutorial: From Install to First Reviewed Edit Cursor Rules Tutorial: Project Rules, User Rules, and AGENTS.md Install Cursor on Windows and Configure a Chinese Interface Cursor MCP Tutorial: Configure, Verify, and Secure MCP Servers Gemini Prompt Guide: Better Instructions and Templates Gemini API Quickstart: Key, Python SDK and First Call Gemini Web App Guide: Login, Files, Chats and Privacy Gemini API Key Security: Storage, Restrictions and Rotation What Is Codex? Capabilities, Limits, and Ways to Use It Codex Beginner Tutorial: Complete Your First Safe Task Install Codex CLI: Sign In and Run Your First Safe Task Codex AGENTS.md Guide: Layered Rules and Validation Codex CLI Commands: Sessions, Review, and Automation How to Use Gemini: Web, Android & iPhone Setup Gemini Features Guide: Chat, Files, Images & Live How to Chat with Gemini: Prompts, Follow-Ups & Live Gemini AI Image Generator Guide: Prompts & Editing Gemini vs GPT-4: Features, Limits & Which to Use Gemini AI Assistant Guide: Mobile, Apps & Privacy Gemini Prompt Engineering Guide: Patterns & Examples Gemini Chat API Guide: Multi-Turn Prompts in Python Gemini System Instructions: API Guide & Examples Gemini Context Caching Guide: Cost, Latency & API
AI Tool Blog Wan3.0 learning hub

Tell a complete story in one thirty-second pass with Wan3.0.

Wan3.0 is Alibaba’s current video model. A single generation can run to 30 seconds — where the model it replaces stopped at 15 — and its reference inputs include documents and webpages, not only text and pictures.

Capabilities and specifications on this page come from Alibaba’s Wan product pages, its Model Studio documentation and the Wan3.0 general-availability announcement. Prices are set by the vendor and change without notice.

Official examples

What Wan3.0 renders

These frames are the stills Alibaba publishes for Wan3.0, taken from the posters of its own demo videos. Between them they cover the ground the release is sold on: simulated light, action choreography, and interiors with real depth.

Official Wan 3.0 example of two armoured figures with glowing blue weapons facing off
Action choreography: two armoured figures with glowing blue weapons facing off, the kind of staged combat the release calls out by name.
Official Wan 3.0 example of a lone desert gas station under a clear blue sky
Simulated daylight: a lone desert gas station under a clear blue sky, where shadows and haze have to agree with the sun.
Official Wan 3.0 example of a caped figure walking with black horses through water
Scale and motion: a caped figure walking with black horses through water, with the animals staying distinct as they move.
Official Wan 3.0 example of a white-haired figure at a fridge in a lit kitchen
Interior depth: a white-haired figure at a fridge in a lit kitchen, the everyday space a product scene needs to look plausible in.
Official Wan 3.0 example of a musician playing a cello in heavy rain
Performance and weather: a musician playing a cello in heavy rain, where the instrument, the hands and the falling water all have to hold.

Example images belong to Alibaba and are reproduced with credit: Wan official site

What it is

How Alibaba describes Wan3.0

Alibaba sells Wan3.0 on one number — thirty seconds in a single pass — and on a new class of input. The vendor’s own announcement is unusually specific about both, including what has not been finished yet.

01

Thirty seconds is the differentiator

Wan2.7 ran 2 to 15 seconds; Wan3.0 generates up to 30 in one pass. That buys room for narrative pacing and continuous camera movement rather than a single shot, and two features make the longer canvas practical: the model recommends a duration from the prompt, and a clip can be extended.

02

Everything to video

Beyond text, images, audio and video, Wan3.0 reads documents and webpages — DOC, XLS, PPT, PDF, TXT, KEY, Pages, Numbers and Markdown, one file or link per request, up to 100 MB and 50 pages. A product deck becomes a launch film; a spreadsheet becomes animated charts.

03

Consistency as a product requirement

Alibaba frames the upgrade around identity, voice and lip sync holding across a sequence, so characters, props, spaces, lighting and style stay aligned through an action sequence or a multi-language performance.

04

The vendor’s own caveat

Alibaba’s general-availability write-up says plainly that audio texture and on-screen text rendering accuracy are still maturing. That is a narrower claim than the marketing pages make, and it is the one to read before committing a deliverable.

API pricing

Per-second pricing by resolution

Alibaba prices Wan3.0 by the second of generated video, at rates that depend only on resolution. These are the standard list prices on the international channel since general availability on 24 August 2026.

Resolution Billing List price
480P Per generated second $0.05
720P Per generated second $0.10
1080P Per generated second $0.20

A 30-second draft at 480P works out at about $1.50 and a 30-second finish at 1080P at about $6.00, before any promotional discount. Alibaba’s China platform lists the same three tiers at ¥0.30, ¥0.60 and ¥1.20 per second.

Capabilities

What Wan3.0 is built to do

These are the capabilities Alibaba lists on the product page, with the vendor’s own wording.

01

Reference to video

Cast referenced characters into new scenes, supporting multi-person and human-object interaction with consistent appearance and voice.

02

Image to video

Alibaba’s wording: generate videos up to 30 seconds with automatic scene splitting, for steadier native audio and picture.

03

Text to video

Cinematic precision with sophisticated motion and highly reliable prompt execution.

04

Video editing

Multi-dimensional, instruction-based editing that reworks visuals, storylines and environments without regenerating the whole concept. Wan2.7 introduced it and Wan3.0 carries it forward.

05

Cinematic control

Lighting, colour grading and composition are tuned to cinematic standards, with complex action choreography and physics simulation called out by name.

06

Immersive audiovisual

Binaural audio with support for regional dialects, cut to the rhythm of the picture so sound and image breathe together.

Specifications

Documented specifications

Every row below is stated in Alibaba’s Model Studio documentation or its general-availability announcement.

Developer
Alibaba
Current model
Wan3.0
Model ID
wan3.0-video
Single generation
Up to 30 seconds
Output resolutions
480P, 720P, 1080P
Duration help
Smart duration recommendation and video extension
Reference inputs
Up to 10 images, 5 video clips and 5 audio tracks
Document inputs
DOC, XLS, PPT, PDF, TXT, KEY, Pages, Numbers, Markdown
Document limits
One file or link per request, up to 100 MB and 50 pages
Audio
Generated with the video, including lip sync
Video editing
Built in; carried forward from Wan2.7
Weights
Not published
How to use it

Documented access channels

Wan3.0 reached general availability on Alibaba Cloud’s international channels on 24 August 2026 and is billed by the second. Alibaba also publishes downloadable weights for the earlier models in the family.

Wan official site

Alibaba’s product page for the model, including the capability list quoted on this page.

Open

Wan3.0 on Model Studio

The model’s documentation on Alibaba Cloud’s platform: input modalities, the three resolutions and the 30-second ceiling.

Open

General-availability announcement

Alibaba’s own write-up of the release, with the pricing table and the caveat about audio and on-screen text.

Open

Open-source releases

Alibaba publishes downloadable weights for the earlier Wan models, up to Wan 2.2 under Apache 2.0. Wan3.0 itself has no published weights, so this repository is the family’s open half, not this model.

Open

Third-party platforms

Wan3.0 is also served through third-party API platforms. The model is the same; the pricing, rate limits and terms are theirs.

Common questions

Wan3.0 questions

How long can a single Wan3.0 generation be?

Up to 30 seconds in one pass, at up to 1080P. Alibaba says the model analyses the prompt and recommends a duration, and that a generation can be extended to continue a storyline.

Does Wan3.0 make its own audio?

Yes. Native audio and picture are produced together, including lip sync. Alibaba’s own general-availability post notes that audio texture is still maturing, so treat the audio as something to review rather than something to ship unchecked.

Can I feed Wan3.0 a document?

Yes, and that is the genuinely new input class. DOC, XLS, PPT, PDF, TXT, KEY, Pages, Numbers and Markdown are accepted alongside webpages — one file or link per request, up to 100 MB and 50 pages.

How many references can I combine?

Alibaba’s own listing says up to 10 images, 5 video clips and 5 audio tracks in one generation, alongside the document or webpage input.

Are Wan3.0’s weights open?

No. Alibaba has not published weights for Wan3.0; it runs through Alibaba Cloud Model Studio and third-party platforms. The downloadable releases in this family stop at Wan 2.2, which shipped under Apache 2.0.

What does the Wan3.0 API cost?

Standard list pricing on the international channel is $0.05 per second at 480P, $0.10 at 720P and $0.20 at 1080P, so a 30-second 1080P generation is about $6.00 before any discount. Alibaba’s China platform lists the same tiers at ¥0.30, ¥0.60 and ¥1.20 per second.

Can Wan3.0 edit a video I already have?

Yes. Wan2.7 introduced video editing and Wan3.0 carries it forward: visuals, plot and dialogue can be modified without regenerating the concept from scratch.