Text-to-video Image-to-video Alibaba Commercial use
wanlogo

WAN 2.5
– Text-to-video
model in Phygital+

WAN 2.5 is an AI video generation model by Alibaba's Tongyi Lab that generates clips up to 10 seconds at resolutions up to 1080p, with automatic native audio synchronization and multilingual dialogue support across English, Russian, Spanish, and more.

Built for video creators, marketers, and developers, it turns briefs into fast, affordable clips with dialogue built in — backed by an open-ish license that supports self-hosted and node-based workflows through community tools like ComfyUI.

wanlogo

At a glance WAN 2.5

Built for

Fast affordable video generation, multilingual dialogue, product and social content

Video and marketing teams that need cost-efficient generation with automatic audio and open-ish licensing

Vendor
Alibaba (Tongyi Lab)
Category
AI video — text to video, image to video
Modalities
Text / images → video with synchronized audio
Commercial use
Yes
Access
All Phygital+ plans
Duration
5–10s per clip; up to 1080p, 24fps
Licensing
Open-ish; self-hosting and ComfyUI supported

Why use WAN 2.5

Fast, affordable, multilingual video generation with audio built in.

1

Automatic native audio

WAN 2.5 automatically populates generated videos with relevant audio effects and speech, synchronized to the visual in one pass. That removes a separate sound-design step for social and marketing clips.

2

Multilingual dialogue understanding

The model understands and generates dialogue in English, Russian, Spanish, and other languages, with what reviewers describe as a deep situational awareness of storytelling — useful for localized campaigns without a separate translation step.

3

Strong prompt adherence at a lower price point

WAN 2.5 offers surprisingly good prompt adherence and native audio for its cost tier, positioning it as a strong all-rounder rather than a budget compromise — even though it isn't the outright quality leader against closed frontier models.

4

Open-ish licensing and node-based workflows

Unlike fully closed models such as Veo and Kling, Wan's more open licensing lets teams self-host (paying for compute rather than per-clip) or build node-based pipelines through community front-ends like ComfyUI — a genuine point of difference in the category.

5

Fast, efficient generation

WAN 2.5 is positioned for faster generation and lower cost relative to rivals like Runway, Pika Labs, and early Veo releases, making it practical for teams that need volume rather than the single best frame.

Built for video and marketing teams

WAN 2.5 fits wherever teams need fast, multilingual, cost-efficient video generation.

Wan
Wan
Wan
Wan
Wan
Wan
Wan
Wan

How to use WAN 2.5

Add a node, describe your scene, and generate an affordable clip with audio.

Wan
Step 1

Choose your input

Add a WAN 2.5 node to the canvas and describe your scene in a text prompt, or upload a still to animate. Set your target aspect ratio (16:9, 9:16, or 1:1) for the placement.

Wan
Step 2

Generate with WAN 2.5

Set duration (5–10 seconds) and resolution. Native audio generates in the same pass. For dialogue-driven clips, write the line in the target language directly in the prompt.

Wan
Step 3

Refine, chain & export

Chain the node with an image model for the starting frame, an upscaler for delivery resolution, or a voice node for additional narration, then export with commercial rights on your plan.

Generated with WAN 2.5

Real generated videos with the exact prompts — copy any prompt and run it yourself.

Wan
Wan
Wan
Wan
Wan
Wan
Wan
Wan
Start Generating!

Pricing and access

One subscription across 30+ AI models — no per-tool credit balances or separate signups. Credit cost per generation is shown live in the node before you run it.

Free

Try WAN 2.5 free

500 weekly credits to test the model and see the output quality. No credit card required.

  • 500 weekly credits
  • Access to 30+ models
  • Personal use
Join as Free
Pro

Professional access

45,000 monthly credits with cheaper per-credit pricing, commercial rights, and video download.

  • 45,000 monthly credits
  • Commercial use license
  • 15% cheaper credits
Join as Pro
Team

Team collaboration

90,000 monthly credits, up to 10 seats, shared workspace, centralized billing, and priority support.

  • 90,000 monthly credits
  • Up to 10 seats
  • Centralized billing
Join as Team
Enterprise

Enterprise scale

210,000 monthly credits, API and integration support, unlimited seats, and dedicated management.

  • API & integration support
  • Unlimited seats
  • Dedicated manager
Join as Enterprise

Technical Specifications

Full specs for the WAN 2.5 model.

Model family
WAN 2.5 (Alibaba Tongyi Lab video line)
Input modalities
Text, images
Duration
5–10 seconds per clip
Frame rate
24fps
Multilingual
English, Russian, Spanish, and more
Native audio
Automatic audio/video synchronization
Vendor
Alibaba (Tongyi Lab)
Output modalities
Video with synchronized audio
Resolution
480p, 720p, 1080p
Aspect ratios
16:9, 9:16, 1:1
Licensing
Open-ish; supports self-hosting, ComfyUI workflows

WAN 2.5 vs Other AI models

How WAN 2.5 compares with other AI video models in the Phygital+ catalog.

Model

WAN 2.5

Automatic native audio, multilingual dialogue, open-ish licensing, cost efficiency

Seedance 2

Multi-subject scene coherence, up to 12 reference inputs

Kling 3.0

Multi-shot storyboard direction, native 4K

Veo 3.1

Native audio with lip sync, scene extension to ~148s

Best for Fast, affordable, multilingual video Directed scenes with interaction Directed multi-shot sequences Ads and long dialogue-driven clips
Strength Automatic audio + multilingual dialogue Native audio, 12 references Multi-shot direction, native 4K Native audio with lip sync
Strength Open-ish licensing, self-hosting, ComfyUI Identity preservation 60fps, 15s single shots Scene extension to ~148s
Strength Lower cost per generation Higher cost per generation Higher cost per generation Higher cost per generation

Use WAN 2.5 Alongside other AI models In Phygital+

Chain image, video, upscaling, and audio into one repeatable workflow.

Wan

Chain with image model

Design the first frame with an image model, then animate it with WAN 2.5 in one workflow.

Wan

Chain with upscaler

Generate with WAN 2.5, then upscale to delivery resolution for social, ads, or broadcast.

Wan

Chain with voice model

Add extra narration on top of the automatic audio with a dedicated voice node.

Wan

Chain with text model

Write the script and shot list with a text model, then feed it straight into video generation.

Start Generating!

Everything you need to know about WAN 2.5 in Phygital+.

FAQ about WAN 2.5

wanlogo
What is WAN 2.5?

WAN 2.5 is Alibaba's AI video generation model from Tongyi Lab, part of the Wan open-ish video family. It generates clips up to 10 seconds at resolutions up to 1080p, with automatic native audio synchronization and multilingual dialogue support across languages including English, Russian, and Spanish.

How much does WAN 2.5 cost in Phygital+?

Usage is credit-based and included in every Phygital+ plan — from the free plan up to Enterprise. The exact credit cost per generation is shown live in the node before you run it, so you can start for free and scale as you go. This is separate from Alibaba's own cloud pricing or third-party host pricing for WAN 2.5.

Can I use WAN 2.5 outputs commercially?

Outputs created on any paid plan (Starter and up) come with commercial-use rights. Note that Wan's own licensing terms vary by version and hosting method — self-hosted deployments should review the specific model license before commercial use. Check the model card for any provider-specific limits.

Who makes WAN 2.5?

WAN 2.5 is made by Alibaba's Tongyi Lab. It's part of the Wan video generation family, Alibaba's flagship line in this category, which has continued to iterate through subsequent versions (Wan 2.6, Wan 2.7) after 2.5's release.

Is WAN 2.5 free to use?

There's no single official "Wan free plan" the way some closed competitors offer a daily credit allowance, because Wan is a model rather than a single hosted product. In practice, free access comes either through self-hosting (paying only for compute) or through third-party hosts that include a small free credit allotment. Phygital+ has an ongoing free tier with 500 weekly credits for running WAN 2.5 as a node.

How long can WAN 2.5 clips be?

Clips run 5 to 10 seconds per generation at up to 1080p resolution. Some third-party hosts have extended this in beta modes, but 5–10 seconds is the standard range for the model.

Does WAN 2.5 generate audio?

Yes — automatic audio/video synchronization is one of its key differentiators, including relevant sound effects and speech populated in the same generation pass, with support for dialogue across multiple languages.

Is WAN 2.5 open source?

Wan carries a more open licensing model than fully closed competitors like Veo and Kling, which is why community tools like ComfyUI support Wan-based node workflows and self-hosting is possible. It's described as "open-ish" rather than fully open — check the specific license terms for your intended use.

Is there an API for WAN 2.5?

Yes. WAN 2.5 is accessible through Alibaba Cloud's DashScope platform and through third-party API providers such as WaveSpeedAI. In Phygital+, you can run it as a node on any plan, and API and integration support is available on the Enterprise plan.

Why use WAN 2.5 inside Phygital+?

Running WAN 2.5 in Phygital+ means no separate hosting setup or API account for one model: it sits on the same canvas as 30+ other AI tools. Chain image generation, video, upscaling, and voice into one repeatable workflow under a single subscription.

Explore more AI models In Phygital+

Browse the full catalog of 30+ AI models available in Phygital+.

FLUX GPT-Image-2 Hailuo 3.0 Ideogram 3 Kling 3.0 Krea AI Luma Midjourney Nano Banana Pro Recraft V4 Reve Runway 4.5 Seedance 2.5 Seedream 5 Veo 3.1 WAN 2.5

Ready to work Faster with WAN 2.5?

Join 100+ teams using Phygital+ – every model in one workspace

Try Phygital+ free
bg