Text-to-image Image editing Alibaba Qwen Commercial use
qwenlogo

Qwen Image
– Text-to-image
model in Phygital+

Qwen Image is Alibaba's open-weight image foundation model: a 20B multimodal diffusion transformer released under Apache 2.0. It renders readable multi-line text inside images in English, Chinese, Japanese, and Korean, which is the one task most diffusion models still get wrong.

In Phygital+ it runs as a single node for both generation and editing. Prompt it from scratch, or connect up to three reference images to guide style, identity, or layout, with camera angle presets and LoRA support built into the same node.

qwenlogo

At a glance Qwen Image

Built for

Posters, banners, packaging, and any layout that needs legible text inside the image

Design and marketing teams working across English, Chinese, Japanese, and Korean

Vendor
Alibaba (Qwen / Tongyi Lab)
Released
August 4, 2025
Category
AI image — text-to-image and reference-guided editing
Modalities
Text + up to 3 reference images → image
Licence
Apache 2.0 (open weights)
Commercial use
Yes
Access
All Phygital+ plans

Why use Qwen Image

Legible multilingual typography, reference-guided editing, and LoRA control in one node.

1

Text that actually reads

Qwen Image was trained for complex text rendering: multi-line layouts, paragraph-level copy, and correct spelling in both alphabetic and logographic scripts. Posters, packaging mockups, and ad banners come out with headlines you can use, instead of gibberish you have to patch in Photoshop.

2

Four scripts, one model

English, Chinese, Japanese, and Korean all render with font detail and layout coherence preserved. For teams shipping the same campaign across several markets, that removes a separate design pass for every visual.

3

Editing in the same node

Connect up to three reference images and describe only what should change. The model holds semantic meaning and visual realism during edits, so identity, style, and untouched areas stay put instead of drifting.

4

Camera angle presets

Horizontal rotation in 45° steps, three distance settings, and vertical tilt from low angle to high angle give you repeatable framing without rewriting the prompt every time.

5

Apache 2.0 open weights

Qwen Image ships under Apache 2.0, one of the most permissive licences among large image models. Commercial use is unrestricted at the model level, which matters if you also plan to fine-tune or self-host.

Built for design and marketing teams

Qwen Image fits wherever a visual has to carry readable text, survive several rounds of edits, or hold the same character and style across a set.

qwen
qwen
qwen
qwen
qwen
qwen
qwen
qwen

How to use Qwen Image

Add the node, write a prompt or connect references, and generate.

qwen
Step 1

Choose your input

Add a QWEN node to the canvas and describe the result in a text prompt. For editing or style control, connect up to three reference images at up to 1664 px per side.

Step 2

Set size and framing

Pick an aspect ratio preset or set width and height directly, from 512 to 1664 px. Use the Angle Settings group for camera rotation, distance, and vertical tilt, and raise Steps toward 30 when you need more detail.

qwen
Step 3

Refine, chain & export

For edits, describe only what should change so the rest of the image stays as it was. Chain the result into an upscaler, a background remover, or a video node, then export with commercial rights on your plan.

Generated with Qwen Image

qwen
qwen
qwen
qwen
qwen
qwen
qwen
qwen
Start Generating!

Pricing and access

One subscription across 30+ AI models, with no per-tool credit balances or separate signups. Credit cost per generation is shown live in the node before you run it.

Free

Try Qwen Image free

500 weekly credits to test the model and see the output quality. No credit card required.

  • 500 weekly credits
  • Access to 30+ models
  • Personal use
Join as Free
Pro

Professional access

45,000 monthly credits with cheaper per-credit pricing, commercial rights, and video download.

  • 45,000 monthly credits
  • Commercial use license
  • 15% cheaper credits
Join as Pro
Team

Team collaboration

90,000 monthly credits, up to 10 seats, shared workspace, centralized billing, and priority support.

  • 90,000 monthly credits
  • Up to 10 seats
  • Centralized billing
Join as Team
Enterprise

Enterprise scale

210,000 monthly credits, API and integration support, unlimited seats, and dedicated management.

  • API & integration support
  • Unlimited seats
  • Dedicated manager
Join as Enterprise

Technical Specifications

Full specs for Qwen Image in Phygital+.

Model family
Qwen-Image
Released
August 4, 2025
Text encoder
Qwen2.5-VL
Input modalities
Text prompt + up to 3 reference images
Resolution
512–1664 px per side
Steps
1–30 (default 20)
Images per run
Up to 2
Camera control
Rotation 0–315°, three distances, tilt −30° to 60°
Vendor
Alibaba (Qwen / Tongyi Lab)
Architecture
20B MMDiT (multimodal diffusion transformer)
Licence
Apache 2.0
Output modalities
Image
Aspect ratios
1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 5:7, 7:5, auto
Prompt following (CFG)
1.0–10.0 (default 3.5)
LoRA
Trained LoRA and preset modifiers supported

Qwen Image vs Other AI models

How Qwen Image compares with other AI image models in the Phygital+ catalog.

Model

Qwen Image

Open-weight 20B model built for legible multilingual text and reference-guided editing

FLUX

General-purpose quality, wide LoRA ecosystem, inpainting and outpainting

Seedream 5

High-resolution photoreal output and fast iteration

Ideogram 3

Typography-focused generation with style references

Best for Text-heavy layouts: posters, packaging, banners General-purpose generation and styling Photoreal product and lifestyle visuals Logos, badges, and typographic posters
Strength Multi-line English, Chinese, Japanese, Korean typography Aesthetic quality and community LoRAs Detail and resolution Text accuracy in Latin scripts
Strength Up to 3 reference images for identity and style Reference and ControlNet nodes Reference-guided generation Style reference codes
Licence Apache 2.0 open weights Varies by variant Proprietary Proprietary

Use Qwen Image Alongside other AI models In Phygital+

Chain generation, editing, upscaling, and video into one repeatable workflow.

Qwen Image

Chain with an upscaler

Generate at 1664 px, then push the result through an upscaling node to reach print or hero-banner resolution.

Qwen Image

Chain with a video model

Use a Qwen Image frame as the first frame for a video node and animate a layout you have already approved.

Qwen Image

Chain with a text model

Draft the headline and body copy with a text node, then feed it into the prompt so typography and wording are decided in one pass.

Qwen Image

Chain with background removal

Cut the subject out of a generated frame and drop it into a composition, a mockup, or a banner template.

Start Generating!

Everything you need to know about Qwen Image in Phygital+.

FAQ about Qwen Image

qwenlogo
What is Qwen Image?

Qwen Image is the first image generation foundation model from Alibaba's Qwen team, released August 4, 2025. It is a 20B-parameter MMDiT (multimodal diffusion transformer) published under the Apache 2.0 licence, built specifically for complex text rendering inside images and for instruction-based editing.

Why is Qwen Image better at text than other models?

Text rendering was a training objective rather than a side effect. The model pairs a Qwen2.5-VL text encoder with the diffusion transformer, so layout and typography are modelled alongside the image itself. On text benchmarks such as LongText-Bench, ChineseWord, and TextCraft it outperforms other leading models, with the widest margin on Chinese.

Which languages does it render inside images?

English, Chinese, Japanese, and Korean render with font detail and layout coherence preserved. The language you prompt in and the language of the text inside the image are set separately, so you can write the brief in one language and request in-image copy in another.

Can I edit an existing image with Qwen Image?

Yes. Connect up to three images to the Start Images input and describe only what should change. Being specific about the edit reduces drift, so parts of the image you did not mention stay as they were.

What resolution does the Phygital+ node support?

Width and height run from 512 to 1664 pixels, with presets for 1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 5:7, and 7:5. For larger output, chain the result into an upscaling node on the same canvas.

Can I use my own LoRA with Qwen Image?

Yes. The node has a My lora field for LoRAs you have trained in Phygital+, with adjustable strength, plus a separate Modifier slot for preset LoRAs such as qwen-fusion.

How much does Qwen Image cost in Phygital+?

Usage is credit-based and included in every plan, from Free up to Enterprise. The credit cost per generation is shown in the node before you run it, so there is no separate signup or credit balance for this one model.

Can I use the outputs commercially?

Outputs created on any paid plan come with commercial-use rights. Qwen Image itself is published under Apache 2.0, which places no additional restriction on commercial use of the model.

Why use Qwen Image inside Phygital+?

Running it in Phygital+ means no separate signup, no GPU to rent, and no per-model credit balance. It sits on the same canvas as 30+ other AI tools, so generation, editing, upscaling, and video can be chained into one repeatable workflow under a single subscription.

Explore more AI models In Phygital+

Browse the full catalog of 30+ AI models available in Phygital+.

Dialogue Creation ElevenLabs v3 FLUX Gemini Omni GPT-Image-2 Hailuo 3.0 Ideogram 3 Kling 3.0 Kling Omni Kling Omni Image Krea AI LTX Video Luma Magnific Upscale Midjourney Nano Banana Pro OmniHuman Qwen Image Recraft V4 Reve Runway 4.5 Seedance 2.5 Seedream 5 Sound Creation Topaz Upscale Upscale Video Veo 3.1 Voice Clone Voice Creation WAN 2.5 Wan Video 2.1

Ready to work Faster with Qwen Image?

Join 100+ teams using Phygital+ – every model in one workspace

Try Phygital+ free
bg