Wan 3.0 Status and the Wan Model You Can Run Now

Wan 3.0 is Alibaba's video model for 30-second single takes with audio generated in the same pass. This page runs it in the generator above and states what its API document actually specifies — resolutions, duration range, reference inputs and price per second — rather than the spec sheets that circulated before it shipped.

Running here nowSpecs from the API documentThirty seconds in one take

Wan 3.0 Status and the Wan Model You Can Run Now

300 creditsSign in to run this one
How to use

From prompt to clip

Four steps on the generator above, and the second is where most clips are decided.

01

Pick text or image to start

Starting from a still fixes what the scene looks like and leaves the model to handle only the motion, which is the more predictable path.

02

Describe the camera and the subject separately

Say what the camera does and what moves inside the frame as two statements. Prompts that merge them tend to produce a still image with drift.

03

Choose length and resolution

Anywhere from 2 to 15 seconds, at 720p or 1080p. Length costs linearly, so a first pass is cheaper short.

04

Generate, then adjust one thing

Change a single element between runs — the move, or the subject, or the lighting. Rewriting the whole prompt makes it impossible to tell what helped.

AI video generator

What is Wan 3.0?

Wan 3.0 is the current generation of Alibaba's Wan video line, launched in August 2026 after a gated beta. It generates up to 30 seconds in a single continuous take at 480p, 720p or 1080p, produces an audio track together with the picture rather than after it, and accepts reference images, reference video, reference audio, documents and links as input alongside the prompt. It is an API model: the weights are not published, which is the one thing about it that has not changed.

Status page

Checked on 2 September 2026: the Wan-AI organisation on HuggingFace and the Wan-Video organisation on GitHub both still top out at the 2.2 family plus the Animate and Dancer branches. It is available through provider APIs; there is no 3.0 checkpoint to download, and any page offering one is not describing this model.

Until it shipped, this page said plainly that it had not, and ran the current release instead. That was accurate on the day it was written and is kept on the record here rather than quietly deleted.

Creative engine

What a clip costs here

Billed per second of output at the rate the model charges, by resolution. Nothing else moves it — audio is included rather than surcharged.

150
Credits, 5s at 480p
300
Credits, 5s at 720p
600
Credits, 5s at 1080p
2-30s
Duration range

What the schema specifies

Resolution is 480P, 720P or 1080P. Duration is an integer from 2 to 30 seconds, or -1 to let the model choose its own length. Audio defaults to on. Aspect ratio covers 16:9, 9:16, 4:3, 3:4, 1:1 and an adaptive mode that picks from the input. Reference images, video, audio, files and links each have their own parameter, as does pinning a first or last frame.

Why this page

One long take, sound included

Most video models hand you four to ten seconds and leave the rest to editing. This one's argument is that the cut is the problem, and it removes it.

01

Thirty seconds without a seam

A single generation covers the whole clip, so light, wardrobe and the way a character moves do not have to be matched across pieces that were never told about each other.

02

Audio arrives with the picture

Sound is generated in the same pass rather than dubbed over the finished video, and it is priced into the per-second rate rather than charged as an extra.

03

More than a prompt to work from

Reference images, a reference clip, reference audio, a first or last frame — each is its own input, so look, movement and rhythm do not have to compete for room in one sentence.

Features

What Wan 3.0 does

Capabilities taken from the model's own API document. Each one is a parameter the model accepts, not a claim about how well it uses it.

Text to video

Generate a clip from a written description alone, up to thirty seconds in one take.

Image to video

Start from a still and describe the motion. Reference images are a first-class input rather than a mode.

Audio in the same pass

Sound is generated together with the picture rather than added over the finished clip, and is on by default.

Reference video and audio

A reference clip supplies motion and camera language; reference audio drives pacing. Both ride alongside the prompt.

Specs

Wan 3.0 specifications

From the model's published API schema and price sheet, read on 2 September 2026.

Release status
Launched August 2026 after a gated beta; available through provider APIs
Open weights
None published — the Wan-AI HuggingFace and Wan-Video GitHub organisations still top out at the 2.2 family
Resolution
480P, 720P or 1080P
Duration
Integer, 2 to 30 seconds in one generation, or -1 for a length the model chooses
Aspect ratio
16:9, 9:16, 4:3, 3:4, 1:1, or adaptive
Audio
Generated with the video, on by default, no separate charge
Reference inputs
Images, video, audio, files and links, plus first-frame and last-frame pinning
Modes wired here
Text to video and image to video
Pricing here
150 credits for a 5-second clip at 480p, 300 at 720p, 600 at 1080p
Use cases

What people make with it

Where a clip with its own soundtrack and a real camera move earns its keep.

01

Social video with native sound

Short vertical clips where the audio is part of the post rather than a track laid underneath it.

02

Establishing and atmosphere shots

Landscape, sky and city moves used to open a sequence or bridge between two pieces of footage.

03

Product and concept motion

A slow move around an object or an idea, where the camera work carries the shot.

04

Storyboard and previz

Roughing out how a sequence reads before anything expensive gets committed to it.

SayMaker pricing — you pay per generation, never per seat

One credit pool covers Nano Banana images and Seedance and Veo video. The cost shows before every run, the safety check runs before the model does, and a failed run is refunded automatically. Use a subscription for ongoing work, or a one-time pack when you just need to top up.

50% OFFYour first purchase is half price
No sign-up

Free

Start generating before you pay anything, and keep earning credits every day you come back. No card, no account needed.
Incl. tax
$0
Start generating

What's included

  • 15 credits before you sign up
  • +40 when you create an account
  • +100 a week from daily check-ins
  • = 488 a month — 97 images or 12 videos
  • No card, no account
  • Free engine every day Grok Imagine
  • Photo editing Nano Banana 2 Lite
  • Not included: Clean downloads, no watermark
  • Not included: Every model unlocked
  • Not included: No shared queue

Secure checkout by Stripe. Card details never touch our servers.

Start here

Starter

A monthly balance that keeps a steady run of clips and images going — on every model SayMaker carries.
Incl. tax
$59.80
$29.90 / month
The whole shelf is open from day one.
Need more each month?1x · $29.90
Total: 3,150 credits per month

What's included

  • 3,150 credits per month
  • Up to 630 images or 78 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • Every model unlocked Veo 3.1 Seedance Nano Banana
  • Cost shown before you generate
  • Generation history

Secure checkout by Stripe. Card details never touch our servers.

Popular

Creator

For regularly making SayMaker drafts, final clips, and the image assets that go with them.
Popular
$119.80
$59.90 / month
The best pick if you create on a regular basis.
Need more each month?1x · $59.90
Total: 7,200 credits per month

What's included

  • 7,200 credits per month
  • Up to 1,440 images or 180 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • Every model unlocked Veo 3.1 Seedance 2 Nano Banana 2
  • No shared queue
  • Priority credit top-ups

Secure checkout by Stripe. Card details never touch our servers.

Scale

Studio

For more output, high-resolution testing, and larger batches.
Incl. tax
$199.80
$99.90 / month
Ideal when several campaigns share one pool.
Need more each month?1x · $99.90
Total: 13,200 credits per month

What's included

  • 13,200 credits per month
  • Up to 2,640 images or 330 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • Every model unlocked Seedance 2.5 Kling 3.0 Veo 3.1
  • No shared queue
  • Built for 4K runs

Secure checkout by Stripe. Card details never touch our servers.

Credit packs

One-time top-ups — buy extra credits any time you run low.

Credit Pack 1500

Pay once, generate whenever. Credits land instantly and never expire.
$39.80
$19.90
Buy more at once1x · $19.90
Total: 1,500 credits

What's included

  • 1,500 credits that never expire
  • Up to 300 images or 37 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • No ongoing subscription
  • Available across the supported models
  • Credits never expire

Secure checkout by Stripe. Card details never touch our servers.

Flexible

Credit Pack 3600

Pay once for a run of final video renders. Credits land instantly and never expire.
$79.80
$39.90
Buy more at once1x · $39.90
Total: 3,600 credits

What's included

  • 3,600 credits that never expire
  • Up to 720 images or 90 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • No ongoing subscription
  • Available across the supported models
  • Use for both video and image generation
  • Credits never expire

Secure checkout by Stripe. Card details never touch our servers.

Batch

Credit Pack 9000

A production-scale top-up for large prompt batches.
$159.80
$79.90
Buy more at once1x · $79.90
Total: 9,000 credits

What's included

  • 9,000 credits that never expire
  • Up to 1,800 images or 225 videos Grok Imagine Nano Banana
  • Clean downloads, no watermark
  • No ongoing subscription
  • Handy for Seedance 2.0 final renders
  • Credits never expire

Secure checkout by Stripe. Card details never touch our servers.

Not sure which to choose? Contact us

Charges appear as “SAYMAKER AI” on your card statement. You can cancel any time from Settings → Billing; cancellation takes effect at the end of the period you already paid for. Operator details, the full model list, and the refund window are on the about page.

Powered by

One credit balance for every SayMaker image and video model

Use one balance across every image and video model on the shelf — the main ones are listed below — and check the credit cost before each request.

Image models

Model credit guide

What each model costs in credits

A failed run is refunded automatically. The exact estimate in the generator varies by model, length, resolution, audio, and number of images.

TypeModelCredit cost
VideoSeedance 2.5The long-take tier, priced per second: 5s 480p ≈ 525 credits, 5s 720p ≈ 1,185. A full 30s take runs ≈ 3,150 at 480p and ≈ 7,090 at 720p.
VideoSeedance 2.0Supplying a starting frame or clip costs less than starting from words alone: 6s 720p ≈ 565 credits from an image, ≈ 925 from text. Scales with resolution and length.
VideoSeedance 2 FastFaster and lower cost. 6s 720p text-to-video ≈ 745 credits.
VideoSeedance 2 MiniThe cheapest tier of the Seedance 2 family. 6s 720p ≈ 465 credits, 6s 480p ≈ 215 credits.
VideoSeedance 1.5 ProAudio doubles the rate. 6s 720p ≈ 85 credits silent, ≈ 160 with audio; 1080p ≈ 175 and ≈ 340.
VideoVeo 3.1Billed per video, not per second, and the tier you pick is the whole price: Lite ≈ 115 credits at 720p and ≈ 135 at 1080p, Fast ≈ 225 at 720p, Quality ≈ 940 at 720p.
VideoKling 3.0Audio raises the rate by about half. 6s 720p ≈ 320 credits silent, ≈ 455 with audio; 1080p ≈ 405 and ≈ 610.
VideoMiniMax H3Fixed 2K, no resolution ladder. Priced per second — a 5s clip ≈ 395 credits.
ImageNano Banana 2Generate or edit from text and images. 20 credits per 1K image, 30 at 2K, 45 at 4K.
ImageNano Banana ProConsistent run times across generations. 30 credits per 1K or 2K image, 55 at 4K.
ImageNano Banana 2 LiteFaster, simpler variant, and the everyday editing price: a flat 15 credits per image at every resolution.
ImageGPT Image 2The lowest-cost premium image model here. 10 credits per 1K image, 15 at 2K, 30 at 4K.
ImageGPT Image 2.5 FlareOpenAI's newest image model, on the fast tier. 25 credits per 1K image, 40 at 2K, 60 at 4K — same price whether you generate or edit.
ImageGPT Image 2.5 SunburstThe 2.5 tier for edits that touch only what you named. Same price as Flare: 25 credits per 1K image, 40 at 2K, 60 at 4K.
ImageSeedream 5.0 LiteFlat pricing — 20 credits per image whether you generate or edit.
FAQ

Wan 3.0 FAQ

What people ask about running Wan 3.0.

The model itself. The model id was probed against the live endpoint before this page was changed, and the parameters the generator sends — resolution, duration, aspect ratio, reference images — are the ones its published schema documents.

Up to thirty seconds in a single generation. The schema takes an integer from 2 to 30, and passing -1 asks the model to choose a length itself. Thirty seconds arrives as one continuous take rather than as clips stitched together, which is the difference that matters when a character has to stay the same character throughout.

480P, 720P or 1080P, in 16:9, 9:16, 4:3, 3:4, 1:1 or an adaptive ratio chosen from the input. Vertical 9:16 comes straight out of the model, so a social cut needs no reframing afterwards, and 1080p holds detail across a full thirty-second take rather than only across a short one.

No. Checked on 2 September 2026, the Wan-AI organisation on HuggingFace and the Wan-Video organisation on GitHub both still stop at the Wan 2.2 family plus the Animate and Dancer branches. It runs through provider APIs, and a download offered as 3.0 weights is not this model.

A five second clip is 150 credits at 480p, 300 at 720p and 600 at 1080p; a full thirty second take is 900, 1,800 or 3,600. Billing is per second of output, so length and resolution are the only things that change the price — audio is included in the rate rather than surcharged.

Yes, and it is on by default. Audio is produced together with the picture rather than added afterwards, which is why the sound tends to match what is happening on screen instead of running alongside it.

Wan ships synchronised audio inside the rate and runs to thirty seconds in one take, which is where it separates from everything else here. Kling is the stronger motion engine — for a dance or an action shot go there; for a long ambient scene with sound included, Wan is the one built for the length.

Make a clip with the current Wan model

Describe the camera move and the subject move as two separate statements, keep the first attempt short, and see what comes back.

Explore other models

Nano Banana 2

HOT

Generate and edit high-quality images from text or photos. 1K, 2K, 4K.

Nano Banana Pro

Consistent run times across generations. Text to image and image editing at 1K, 2K, 4K.

Seedance 2.5

NEW

The newest Seedance tier. Text or photo to video at 480p, 720p or 1080p, priced per second.

Seedance 2.0

Generate video from text or images. 480p–1080p, with audio.

Veo 3.1

Cinematic video. Lite and Fast tiers, 720p–4K.

Kling 3.0

Text or photo to video at 720p, 1080p or 4K, audio optional. Priced per second.

Seedance 1.5 Pro

Mid-weight text or image to video. 480p–1080p, audio optional, priced per second.

MiniMax H3

NEW

Text or photo to video at a fixed 2K, 5 to 10 seconds. Priced per second, no tier to pick.

GPT Image 2

Text to image and description-based edits. 1K–4K, priced per image.

Seedream 5.0 Lite

Light-tier image generation and edits at one flat credit cost.

Nano Banana 2 Lite

The free image editor. Text to image and photo edits at one flat credit cost.

Nano Banana

The original Nano Banana. Text to image and single-photo edits, priced per image.

Gemini Flash Image

Google's fast Gemini image line. Text to image and multi-photo edits, 1K–4K.

Gemini 3 Pro Image

Google's top Gemini image tier. Sharper type and detail, 1K–4K.

ChatGPT Image Generator

OpenAI's image model. Strong in-image text and layout, 1K–4K.

GPT Image 2.5

OpenAI's newest image model, released 8 Sep 2026. Both tiers run here: Flare for speed, Sunburst for edits that stay put.

Seedance 2 Mini

The cheapest Seedance 2 tier. 480p or 720p video, priced per second.

Seedance 2 Fast

The middle Seedance 2 tier. Video from text or a still image at 480p or 720p, priced per second.

Kling 2.6

Text or photo to video, 5 or 10 seconds, with generated sound optional.

Kling 3.0 Turbo

NEW

The fast Kling 3.0 tier. Text or photo to video at 720p or 1080p, priced per second.

Seedream 4.0

The entry Seedream tier. Text to image at 1K, 2K or 4K across nine framing presets.

Seedream 4.5

Photoreal text to image at eight aspect ratios, with a basic and a high quality setting.

Seedream 5.0 Pro

The top Seedream tier. Photoreal text to image at eight aspect ratios, basic or high fidelity.

Qwen Image 3

NEW

Built for readable text and page layout. Posters, infographics and storyboards at 1K or 2K.

LTX 2.5

NEW

Open-weights video with synchronised audio — up to 4K, 20 seconds, 50fps.

Grok Imagine Image 2.0

NEW

xAI's Quality Mode model — designed typography, layouts that hold, and up to five reference images per run.