Wan AI

Wan 3.0

Alibaba's all-in-one video renderer — 30-second single-pass clips, native audio, and one reference set that takes images, video, and voice together.

Brak subskrypcji
Kredyty nigdy nie wygasają
Dowiedz się więcej

Zapłać raz za kredyty - używaj ich w każdym modelu na ZOOOP. · Doładuj, kiedy potrzebujesz, bez miesięcznego spalania.

Powered by Wan AI's API on ZOOOP

Kluczowe cechy

30 seconds in a single pass

One request produces up to 30 seconds of continuous video — long enough for a full narrative beat, a one-take walkthrough, or a complete ad read, without stitching clips together afterwards.

All-in-One Reference

Up to ten reference images, five reference clips and five reference audio tracks go into one request together. Faces, products and voices stay aligned across the whole shot instead of drifting between takes.

First and last frame control

Lock the opening frame, optionally lock the closing frame, and Wan generates the motion that bridges them — the most reliable way to land an exact ending for a cut.

Native synchronized audio

Dialogue, ambient sound and music are generated with the picture in the same pass and arrive as a stereo track on the finished file. Toggling audio off costs nothing extra.

Przypadki użycia

One-take 30-second stories

One-take 30-second stories

A single generation covers a complete narrative beat — a continuous camera move, a full ad read, a walkthrough — instead of three clips that have to be matched in an editor.

Character consistency across a shot

Character consistency across a shot

Feed up to ten reference images of the same person, costume or product and the identity holds through the whole clip — the same face at second 2 and second 28.

Voice-referenced dialogue scenes

Voice-referenced dialogue scenes

Reference audio goes into the same request as the visual references, so a spoken scene lands with the voice you supplied rather than a generic synthetic read.

Product reveals with a locked ending

Product reveals with a locked ending

Start on the packshot, end on the hero frame. First/last frame control fixes both ends so the reveal cuts cleanly into whatever follows it.

1440p hero deliverables

1440p hero deliverables

The top tier renders at 2560×1440 for landing-page heroes and large-screen playback, where a 720p master visibly softens.

Fast iteration on the Prime tier

Fast iteration on the Prime tier

Wan 3.0 Prime is the same model on an express lane — identical controls, prioritized generation — for the pass where you are still deciding the shot.

Wybierz odpowiedni model

Wan 3.0 is the pick when a shot has to run long and stay consistent with real reference material. Switch when the job wants something else.

30-second single-pass shots with mixed referencesWan 3.0
Open weights and instruction-based editsWan 2.7
Multi-modal reference with beat-aware audioSeedance 2.0
Reference images, clips and voice in one video callMiniMax H3
Top-end fidelity and a 4K delivery pathVeo 3.1
Cheapest fast drafts on the same familyWan 2.6 Flash

Jak używać

01

Open Wan 3.0 from this page or pick it in the Video Generator.

02

Write the prompt, and add reference images, clips or audio if the shot needs them.

03

Pick aspect ratio, resolution up to 1440p, and a duration between 2 and 30 seconds.

04

Generate — the finished clip arrives with its audio track already synced.

Głębokie nurkowanie

What Wan 3.0 is good at — and what it's not

Wan 3.0 is Alibaba's answer to the question every video model has been dodging: what happens when the shot needs to last longer than a few seconds? Most flagships cap a single generation somewhere between five and fifteen seconds, which means anything resembling a scene gets assembled in an editor from clips that never quite match — the face drifts, the light shifts, the room changes shape at the cut. Wan 3.0 renders up to thirty seconds in one pass, and that single change reorganizes how you plan a shot. A continuous camera move through a space, a complete thirty-second ad read, a one-take product walkthrough: these stop being edit problems and become prompt problems.

The second thing that sets it apart is All-in-One Reference. Reference systems usually take one kind of input — some models take images, a few take a clip, a rare one takes a voice. Wan 3.0 takes all three in the same request: up to ten reference images, five reference clips and five reference audio tracks, capped at fifteen seconds of reference video and fifteen seconds of reference audio. In practice that means a whole brand kit or a whole character sheet goes in at once, and the identity holds for the length of the clip rather than for the first few seconds. On ZOOOP the routing is automatic — attach any reference and the request goes to the reference pipeline; leave them empty and the same model runs plain text-to-video.

Two more capabilities matter day to day. First and last frame control lets you pin both ends of a clip and have Wan generate the motion between them, which is how you land an exact ending that cuts cleanly into the next shot. And native synchronized audio — dialogue, ambience, music — is generated with the picture in the same pass and arrives as a stereo track on the file, with the audio toggle costing nothing either way.

Where it's weaker: it is not open-weight. Wan 2.7 shipped under Apache 2.0 with published weights, and Wan 3.0 did not — no weights, API only. For a pipeline that needs self-hosting or open provenance, that is a hard stop and 2.7 is still the answer. Its 1440p tier is an enhanced upscale, not a native render — a real 2560×1440 file, good for large-screen delivery, but a model with a native high-resolution pass will hold detail better. There is no negative prompt, so exclusions have to be handled by describing what you do want. And on image-to-video the output follows the aspect ratio of your input frame — there is no ratio control on that path, so crop the source before you generate.

A reasonable mental model: reach for Wan 3.0 when the shot is long, or when consistency against real reference material is the thing that decides whether the take is usable. For open weights and instruction-based edits, Wan 2.7. For top-end single-shot fidelity, Veo 3.1. For quick throwaway drafts on the same family, Wan 2.6 Flash. And when you are still deciding what the shot is, run the Prime tier and switch back for the final render.

Najczęściej zadawane pytania

How long can one Wan 3.0 clip be?+

Between 2 and 30 seconds in a single generation. That is the headline difference from the previous generation, which topped out well short of that — a 30-second beat that used to need three clips and an edit is now one request.

What can go into the reference set?+

Up to ten reference images, five reference clips and five reference audio tracks, all in the same request. Reference clips and reference audio are each capped at 15 seconds in total. Adding any reference automatically routes the request to the reference pipeline; a text-only prompt runs text-to-video.

Is Wan 3.0 open source like Wan 2.7?+

No. Wan 2.7 shipped under Apache 2.0 with open weights; Wan 3.0 is API-only, with no published weights. If open-weight provenance or self-hosting matters for your pipeline, Wan 2.7 remains the model to use.

Is 1440p a native render?+

The 480p, 720p and 1080p tiers are native renders. The 1440p tier is produced through an enhanced upscale of a native render rather than a native 1440p pass — it delivers a genuine 2560×1440 file, and it is the right pick for large-screen delivery, but it is not the same thing as a native 1440p model.

What is the difference between Wan 3.0 and Wan 3.0 Prime?+

Same model, same parameters, same limits — Prime runs on an express lane for quicker turnaround at a higher rate. Use the standard tier for final renders and Prime when you are iterating and waiting is the expensive part.

Does Wan 3.0 support negative prompts?+

No. Wan 3.0 does not take a negative prompt — describe what you want in the positive prompt instead. Being specific about camera angle, lighting and motion does more here than trying to exclude things.

Więcej modeli

Wan AI
Wan V2.7
Wan AI
ByteDance
Seedance V2.0
ByteDance
MiniMax
MiniMax H3
MiniMax
Google
Veo 3.1
Google
OpenAI
GPT Image 2.5
OpenAI
OpenAI
GPT Image 2.0
OpenAI
ByteDance
Seedance V2.0 Fast
ByteDance
ByteDance
Seedance V1.5 Pro
ByteDance
ByteDance
Seedance V1.0 Pro
ByteDance
ByteDance
Seedance V1.0 Pro Fast
ByteDance
ByteDance
Seedance V1.0 Lite
ByteDance
ByteDance
Seedream 5.0 Pro
ByteDance
ByteDance
Seedream 5.0 Lite
ByteDance
ByteDance
Seedream 4.5
ByteDance
ByteDance
Seedream 4
ByteDance
ByteDance
Dreamactor V2
ByteDance
MiniMax
Minimax Music V2.6
MiniMax
MiniMax
Minimax Music V2
MiniMax
MiniMax
Speech-2.8-HD
MiniMax
MiniMax
Speech-2.8-Turbo
MiniMax
Kling AI
Kling O3
Kling AI
Kling AI
Kling V3
Kling AI
Kling AI
Kling V3 Pro
Kling AI
Kling AI
Kling V2.6 Pro
Kling AI
Kling AI
Kling V2.6
Kling AI
Kling AI
Kling Lipsync
Kling AI
Kling AI
Kling Avatar V2
Kling AI
Kling AI
Kling O1
Kling AI
Midjourney
Midjourney V8.2
Midjourney
Midjourney
Midjourney V8.1
Midjourney
Midjourney
Midjourney
Midjourney
Alibaba
Happy Horse V1.1
Alibaba
Alibaba
Happy Horse
Alibaba
xAI
Grok Imagine V1.5
xAI
xAI
Grok Imagine
xAI
xAI
xAI TTS
xAI
Google
Veo 3.1 Fast
Google
Google
Veo 3
Google
Google
Nano Banana Pro
Google
Google
Nano Banana 2
Google
Google
Nano Banana
Google
Google
Lyria 3 Pro
Google
Google
Lyria 3
Google
Google
Lyria2
Google
Google
Gemini 3.1 Flash TTS
Google
Wan AI
Wan V2.2
Wan AI
Wan AI
Wan V2.2 Turbo
Wan AI
Wan AI
Wan V2.5
Wan AI
Wan AI
Wan V2.6
Wan AI
Wan AI
Wan V2.6 Flash
Wan AI
Pixverse AI
Pixverse V6
Pixverse AI
Pixverse AI
Pixverse V5.5
Pixverse AI
Pixverse AI
Pixverse V5
Pixverse AI
Pixverse AI
Pixverse Lipsync
Pixverse AI
Vidu AI
Vidu Q3 Pro
Vidu AI
Vidu AI
Vidu Q3
Vidu AI
Vidu AI
Vidu Q3 Turbo
Vidu AI
Vidu AI
Vidu Q2 Pro
Vidu AI
Vidu AI
Vidu Q2 Turbo
Vidu AI
Luma AI
Luma Ray 2
Luma AI
Luma AI
Luma Ray 2 Flash
Luma AI
Flux AI
Flux 2 Pro
Flux AI
Flux AI
Flux 2
Flux AI
Flux AI
Flux 2 Flash
Flux AI
ElevenLabs
Multilingual V3
ElevenLabs
ElevenLabs
Multilingual V2
ElevenLabs
ElevenLabs
Sound Effects V2
ElevenLabs
Hailuo AI
Hailuo 2.3
Hailuo AI
Hailuo AI
Hailuo 2.3 Fast
Hailuo AI
Hailuo AI
Hailuo 02
Hailuo AI
Lightricks
LTX-2.3 Pro
Lightricks
Lightricks
LTX-2.3 Fast
Lightricks
Lightricks
LTX-2.3
Lightricks
Pika AI
Pika V2.2
Pika AI
Qwen
Qwen3-TTS
Qwen
Inworld
Inworld TTS
Inworld
Bilibili Index
Index TTS 2
Bilibili Index
Resemble AI
Chatterbox TTS Multilingual
Resemble AI
Open Source
ACE-Step
Open Source
Open Source
LUX TTS
Open Source
CassetteAI
Music Generator
CassetteAI
Recraft
Text to Vector V4.1
Recraft
Recraft
Text to Vector V4.1 Pro
Recraft
Recraft
Image to Vector
Recraft