Length is set in whole seconds anywhere from 2 to 30. Half a minute gives a scene time to breathe: a chef plating a dish from first ingredient to garnish, a camera drifting through a shop, a character walking into frame and reacting, all inside one generation instead of several stitched together.
IAllegro
Wan 3.0 Video Generator for Long Takes and Reference Images
Wan 3.0, part of Alibaba's Wan family of video models, is the one to open when a scene needs room to unfold, or when the same person or product should stay recognizable from the first second to the last. Start from a text prompt, a first frame or up to 10 reference images, set any length from 2 to 30 seconds, and run it in the generator just below.
Current modelWan 3.0
Andante
Wan 3.0 specs and limits
The settings the generator accepts for this model, gathered in one list so you can plan a shot before you spend any credits.
- Start modes
- Text to video; first frame to video, with an optional last frame; reference images to video
- Clip length
- Any whole number of seconds from 2 to 30
- Resolution
- 480P, 720P or 1080P
- Aspect ratio
- Adaptive (follows your input), 16:9, 9:16, 1:1, 4:3 or 3:4
- Reference images
- Up to 10, addressed in the prompt as Image1, Image2 and so on, in the order you upload them
- Audio
- Generated together with the picture; the toggle is on by default and can be switched off
- Prompt length
- Up to 5,000 characters
- Credits
- The cost updates as you change length and resolution and is shown on the Generate button; failed runs are refunded automatically
- Account and storage
- Sign-in with a Sonata AI account is required; finished clips are kept in History
Adagio
Made for longer takes and recurring subjects
What sets this generator apart comes down to two things: how long a single clip can run, and how much visual material you can hand the model before it starts.
Scherzo
Making a clip in the generator above
Six steps take you from an empty form to a finished file. The settings matter more than the order, but this is the quickest route through them.
Sign in
Log in with your Sonata AI account. Each generation is paid for with credits, so check that your balance covers the clip you have in mind; the form shows the price before anything is spent.
Pick how the clip should start
Choose text to video when all you have is an idea, first frame when the opening image already exists, or reference images when several specific pictures need to show up in the scene.
Add images if the mode needs them
For first frame, upload the opening still and, if you like, a closing still. For references, upload up to ten pictures in the order you plan to mention them, because that order decides which one is Image1, which is Image2 and so on.
Write the prompt
Describe the subject, the action, the camera and the light. The box takes up to 5,000 characters, which is room enough to lay out a sequence of moments for a long clip. Use the image names whenever you point at a reference.
Set length, resolution, ratio and audio
Pick a length between 2 and 30 seconds, then 480P, 720P or 1080P, then adaptive or a fixed ratio, and decide whether the audio toggle stays on. Keep an eye on the Generate button: the credit figure moves as you change length and resolution.
Generate, then check History
Press Generate and wait for the job to finish. The video is saved to History, so you can come back to it later. If the job fails, the credits return to your balance automatically and you can try again.
Adagio
Where long takes and references pay off
A few kinds of projects in which the 30-second limit and the reference images make a practical difference.
Adagio
Prompting long clips and multiple references
A 30-second shot and a board of ten images both reward a little planning. These habits make the prompt easier for the model to follow and easier for you to fix.
Coda
Wan 3.0 questions and answers
Start modes, references, cost, sound and your account, answered briefly.
01Which start mode should I use?
Text to video works when all you have is an idea. First frame suits a shot whose opening image you already have, optionally with a closing image as well. Reference images are for scenes where particular people, products or places must appear; you can upload up to 10 and mention each one by name.
02How do I point at one particular reference image in my prompt?
Use its name. Pictures are numbered in the order you upload them, so the first becomes Image1, the second Image2, and so on up to Image10. A line such as 'the woman in Image1 picks up the mug from Image2' is clearer than describing each picture all over again.
03What decides the credit cost of a clip?
Length and resolution. The figure on the Generate button updates as you change either one, so a 30-second clip at 1080P costs more than a 5-second draft at 480P. You always see the number before you start.
04What happens to my credits if a generation fails?
They are refunded automatically. There is no request to file: the amount goes back to your balance and you can try again, perhaps with a simpler prompt or fewer references.
05Can I make a video without sound?
Yes. Audio is generated along with the picture, and the toggle is on by default. Switch it off before generating if you plan to add your own music, narration or sound design afterwards.
06Should I choose adaptive or a fixed aspect ratio?
Adaptive follows your input, which is handy when an uploaded image already has the framing you want. Pick a fixed ratio, 16:9, 9:16, 1:1, 4:3 or 3:4, when the clip has to fit a particular screen or feed.
07How is this page different from the Wan 2.6 and Wan 2.5 pages?
They are separate generators for separate models. Wan 2.6 on this site makes clips of up to 15 seconds at up to 1080P. Here the limits are different: clips can run to 30 seconds, you can add up to 10 reference images, and the sound has an on/off switch. Choose whichever fits the job.
08What are the input limits?
The prompt takes up to 5,000 characters. In first-frame mode you add one opening image and, if you want, one closing image; in reference mode you can upload up to 10 pictures. Text to video needs no upload at all, just the prompt.
09Where do I find my finished videos?
Every result is saved in History in your Sonata AI account. That keeps a 480P draft and the final 1080P render of the same idea in one place, so you can compare them before you decide what to publish.
10Why did my clip fail or turn out differently than planned?
When a job fails, the credits come back automatically, so trying again costs nothing extra. When a clip finishes but misses the mark, check the prompt first: a reference mentioned under the wrong number, two actions asked for at the same moment, or a very short prompt spread over a long duration.
11Is Sonata AI connected to Alibaba?
No. Sonata AI is an independent platform and is not affiliated with, endorsed by or sponsored by Alibaba. Wan is a trademark of its owner; the name appears here only to say which model the generator runs.