IAllegro

Wan 3.0 Video Generator for Long Takes and Reference Images

Wan 3.0, part of Alibaba's Wan family of video models, is the one to open when a scene needs room to unfold, or when the same person or product should stay recognizable from the first second to the last. Start from a text prompt, a first frame or up to 10 reference images, set any length from 2 to 30 seconds, and run it in the generator just below.

Current modelWan 3.0
Start from
0 / 5000
5 s
Resolution
Aspect ratio

Dialogue, music and ambience are generated together with the picture.

The price updates as you change the settings. If a generation fails, the credits are refunded automatically.

Preview
No video available. Please generate a video first!
Andante

Wan 3.0 specs and limits

The settings the generator accepts for this model, gathered in one list so you can plan a shot before you spend any credits.

Start modes
Text to video; first frame to video, with an optional last frame; reference images to video
Clip length
Any whole number of seconds from 2 to 30
Resolution
480P, 720P or 1080P
Aspect ratio
Adaptive (follows your input), 16:9, 9:16, 1:1, 4:3 or 3:4
Reference images
Up to 10, addressed in the prompt as Image1, Image2 and so on, in the order you upload them
Audio
Generated together with the picture; the toggle is on by default and can be switched off
Prompt length
Up to 5,000 characters
Credits
The cost updates as you change length and resolution and is shown on the Generate button; failed runs are refunded automatically
Account and storage
Sign-in with a Sonata AI account is required; finished clips are kept in History
Adagio

Made for longer takes and recurring subjects

What sets this generator apart comes down to two things: how long a single clip can run, and how much visual material you can hand the model before it starts.

Scherzo

Making a clip in the generator above

Six steps take you from an empty form to a finished file. The settings matter more than the order, but this is the quickest route through them.

  1. Sign in

    Log in with your Sonata AI account. Each generation is paid for with credits, so check that your balance covers the clip you have in mind; the form shows the price before anything is spent.

  2. Pick how the clip should start

    Choose text to video when all you have is an idea, first frame when the opening image already exists, or reference images when several specific pictures need to show up in the scene.

  3. Add images if the mode needs them

    For first frame, upload the opening still and, if you like, a closing still. For references, upload up to ten pictures in the order you plan to mention them, because that order decides which one is Image1, which is Image2 and so on.

  4. Write the prompt

    Describe the subject, the action, the camera and the light. The box takes up to 5,000 characters, which is room enough to lay out a sequence of moments for a long clip. Use the image names whenever you point at a reference.

  5. Set length, resolution, ratio and audio

    Pick a length between 2 and 30 seconds, then 480P, 720P or 1080P, then adaptive or a fixed ratio, and decide whether the audio toggle stays on. Keep an eye on the Generate button: the credit figure moves as you change length and resolution.

  6. Generate, then check History

    Press Generate and wait for the job to finish. The video is saved to History, so you can come back to it later. If the job fails, the credits return to your balance automatically and you can try again.

Adagio

Where long takes and references pay off

A few kinds of projects in which the 30-second limit and the reference images make a practical difference.

Adagio

Prompting long clips and multiple references

A 30-second shot and a board of ten images both reward a little planning. These habits make the prompt easier for the model to follow and easier for you to fix.

Coda

Wan 3.0 questions and answers

Start modes, references, cost, sound and your account, answered briefly.

01

Which start mode should I use?

Text to video works when all you have is an idea. First frame suits a shot whose opening image you already have, optionally with a closing image as well. Reference images are for scenes where particular people, products or places must appear; you can upload up to 10 and mention each one by name.

02

How do I point at one particular reference image in my prompt?

Use its name. Pictures are numbered in the order you upload them, so the first becomes Image1, the second Image2, and so on up to Image10. A line such as 'the woman in Image1 picks up the mug from Image2' is clearer than describing each picture all over again.

03

What decides the credit cost of a clip?

Length and resolution. The figure on the Generate button updates as you change either one, so a 30-second clip at 1080P costs more than a 5-second draft at 480P. You always see the number before you start.

04

What happens to my credits if a generation fails?

They are refunded automatically. There is no request to file: the amount goes back to your balance and you can try again, perhaps with a simpler prompt or fewer references.

05

Can I make a video without sound?

Yes. Audio is generated along with the picture, and the toggle is on by default. Switch it off before generating if you plan to add your own music, narration or sound design afterwards.

06

Should I choose adaptive or a fixed aspect ratio?

Adaptive follows your input, which is handy when an uploaded image already has the framing you want. Pick a fixed ratio, 16:9, 9:16, 1:1, 4:3 or 3:4, when the clip has to fit a particular screen or feed.

07

How is this page different from the Wan 2.6 and Wan 2.5 pages?

They are separate generators for separate models. Wan 2.6 on this site makes clips of up to 15 seconds at up to 1080P. Here the limits are different: clips can run to 30 seconds, you can add up to 10 reference images, and the sound has an on/off switch. Choose whichever fits the job.

08

What are the input limits?

The prompt takes up to 5,000 characters. In first-frame mode you add one opening image and, if you want, one closing image; in reference mode you can upload up to 10 pictures. Text to video needs no upload at all, just the prompt.

09

Where do I find my finished videos?

Every result is saved in History in your Sonata AI account. That keeps a 480P draft and the final 1080P render of the same idea in one place, so you can compare them before you decide what to publish.

10

Why did my clip fail or turn out differently than planned?

When a job fails, the credits come back automatically, so trying again costs nothing extra. When a clip finishes but misses the mark, check the prompt first: a reference mentioned under the wrong number, two actions asked for at the same moment, or a very short prompt spread over a long duration.

11

Is Sonata AI connected to Alibaba?

No. Sonata AI is an independent platform and is not affiliated with, endorsed by or sponsored by Alibaba. Wan is a trademark of its owner; the name appears here only to say which model the generator runs.