Leave the reference area empty and the model draws from your description alone. Drop in a few photos and the same prompt box becomes an editing instruction, so an idea that began as one sentence can later pull in a real product shot.
IAllegro
Create and Edit Pictures with Qwen Image 2.1
Qwen Image 2.1 is the image model from Alibaba's Qwen team, and on Sonata AI it handles two jobs in one panel: drawing a picture from a written prompt, or reworking up to ten reference photos with a single instruction. Write in whatever language you think in, choose 1K or 2K, and start right in the generator below.
Your image will appear here
Andante
Qwen Image 2.1 at a Glance
Every option the generator exposes for this model, along with its limits.
- Developer
- Alibaba's Qwen team
- Ways to start
- Text to image, or editing with reference images
- Reference images
- 1 to 10 per run; the upload order is how the prompt refers to them
- Resolution
- 1K or 2K; 2K holds four times the pixels and takes noticeably longer
- Aspect ratios
- 1:1, 4:3, 3:4, 3:2, 2:3, 16:9, 9:16, 21:9 and 9:21
- Auto ratio
- Edit mode only; the output takes the shape of the first reference image
- Prompt
- Any language, up to 5,000 characters
- Output file
- PNG
- Credits
- Cost depends on resolution and is shown before you generate; failed runs are refunded automatically
Adagio
One Panel for Writing and Editing
What changes when text-to-image and multi-reference editing sit side by side.
Scherzo
How to Use the Generator
Six steps from an empty prompt box to a finished PNG in your history.
Sign in to your account
The generator runs on credits, so sign in to Sonata AI first. Your account holds the credit balance and lists every finished picture in your history.
Choose text or references
Leave the upload slots empty for pure text-to-image, or add one to ten pictures to edit or combine. Upload them in the order you plan to mention them, because that order is how the prompt will point to each one.
Write the instruction
Describe the scene or the change in the language you are most comfortable with. With references, call them by position, such as “the first image” or “the second image”, and say what each should contribute.
Set the frame and resolution
Pick one of the nine aspect ratios, or auto in edit mode to follow the first reference. Then choose 1K or 2K; the credit figure on the button updates, so you know the cost before anything starts.
Generate and wait
Click generate. A 1K picture comes back sooner, while 2K takes noticeably longer because it renders four times as many pixels, which makes it the setting for the version you mean to keep.
Open the result in your history
Finished images land in the history below the generator as PNG files. Download the one you like, or copy its wording into a new run and change one detail to explore a variation.
Adagio
Projects That Suit This Model
Situations where combining references, choosing the frame and writing in your own language pay off.
Adagio
Prompting by Order and by Language
Small habits that make references easier to steer and lettering easier to read.
Coda
Qwen Image 2.1 Questions, Answered
Short answers about references, languages, resolution, credits and results.
01Is Sonata AI connected to Alibaba or the Qwen team?
No. Sonata AI is an independent platform and is not affiliated with, endorsed by or sponsored by Alibaba. Qwen is a trademark of its owner; the name appears here only so you know which model you are selecting.
02How does this page differ from Qwen Image Edit?
Qwen Image Edit is the older single-image editor and keeps its own page. Here you can start from text alone or from one to ten reference images in the same panel, and choose 1K or 2K output. Pick whichever fits the task.
03How do I tell the model which reference is which?
By order. The first picture you upload is “the first image”, the next is “the second image”, up to the tenth. Use those phrases in the prompt, and if you change the upload order, update the wording so it still points to the right pictures.
04Do I have to write prompts in English?
No. Prompts work in any language, up to 5,000 characters. Where lettering matters, Chinese and English both work well, so you could describe a market scene in German and quote the Chinese characters for a shop sign.
05When is 2K worth the extra wait?
2K has four times the pixels of 1K and takes noticeably longer, and its credit cost differs. It pays off for prints, large displays, tight crops and small lettering. For testing a composition, 1K is usually enough.
06What does the auto aspect ratio do?
Auto appears only when you edit with reference images. It gives the result the same shape as the first reference, which helps when that photo already has the framing you want. In text-to-image mode, pick one of the nine fixed ratios.
07What does a generation cost, and what if it fails?
The credit cost depends on the resolution you choose, and the exact figure is shown on the button before you generate. If a generation fails, the credits go back to your balance automatically, with nothing to request.
08Why did my edit use the wrong subject or frame?
Check the order first. If the photo you call “the first image” was not actually uploaded first, the model follows the upload order, not your intention. With auto selected, the frame copies the first reference, so move the picture whose shape you want to the front.
09Do I need an account, and where do results go?
Yes, generating requires signing in, because every run uses credits from your balance. Each result is a PNG and appears in the history under the generator, where you can download the pictures you want to keep.