Logo
Video Models

WAN 3.0 Text / Image / Reference to Video Generator

Generate video with WAN 3.0, Alibaba Tongyi Lab's unified video model, on Dreamega AI. Write a prompt for text-to-video, animate a first frame with an optional last frame for image-to-video, or combine up to 10 reference images, 5 reference videos and 5 reference audio tracks for reference-to-video. Every mode renders 2 to 30 seconds at 480p, 720p or 1080p with native synchronized audio.

Public
*

WAN 3.0 Video Gallery

Browse clips generated with WAN 3.0 from text prompts. Every example runs in a single pass with native synchronized audio.

Create with WAN 3.0
AI Video

Astronaut in the Overgrown City

A lone astronaut crosses a reclaimed city and finds a child holding a glowing flower, generated from text.

Prompt

A lone astronaut walks through the ruins of a once-busy city, surrounded by abandoned cars, overgrown skyscrapers, and trees growing through cracked streets. She discovers a small child standing inside an old convenience store, holding a glowing flower. The astronaut slowly removes her helmet as birds suddenly rise into the sky and sunlight breaks through the clouds. Emotional science-fiction film, grand post-apocalyptic environment, slow cinematic camera movement, wide establishing shots, intimate facial close-ups, realistic dust particles, warm sunlight contrasting with cold ruins, epic yet hopeful atmosphere.

Live PipelineTake 01 / 01

WAN 3.0 Image to Video Gallery

See how WAN 3.0 brings a still image to life. The clip below starts from a single first-frame image and renders with native synchronized audio.

Source Feeds01 Inputs
Ranger Meets the Forest Dragon - Input 1
Program · On AirAI · Generated
Output
Transcript · 01

Ranger Meets the Forest Dragon

The dragon slowly opens its eyes, leaves move from its breathing, glowing particles float around the forest, the ranger slowly steps forward and reaches out a hand. The camera slowly circles around both characters revealing the enormous scale difference.

WAN 3.0 YouTube Videos

Watch community demos, reviews and tutorials showcasing what WAN 3.0 can do

  • Wan 3.0 Is Here — 30-Second AI Videos in One Pass - Tech With Hamza
  • China Did It AGAIN? – 100+ Wan 3.0 AI Videos - Airt
  • WAN 3.0 Is Almost Too Good to Be True.. (Review) - Oprelia AI
  • Wan 3.0 Public Beta Just Launched and is FREE TO TRY! - ByteForward
  • How to Access Wan 3.0 API - Best Alternative to Seedance 2 - Anil Chandra Naidu Matcha

WAN 3.0 YouTube Videos

Watch community demos, reviews and tutorials showcasing what WAN 3.0 can do

Reel · Specifications

What's WAN 3.0

Alibaba Tongyi Lab's unified video model, served on Dreamega AI

  1. · 012-30sSingle-Pass Duration
  2. · 021080pMax Resolution
  3. · 033Modes: Text / Image / Reference
  4. · 0410Reference Images per Run

WAN 3.0 is Alibaba Tongyi Lab's unified video generation model. Where WAN 2.7 split text-to-video, image-to-video, reference generation and editing across separate models, WAN 3.0 folds them into one. On Dreamega AI it is available in three modes: Text-to-Video builds a clip from a written prompt, Image-to-Video animates a first-frame image with an optional last frame, and Reference-to-Video combines up to 10 reference images, 5 reference videos and 5 reference audio tracks to keep a subject consistent. Every mode renders 2 to 30 seconds in a single pass at 480p, 720p or 1080p, with native synchronized audio on by default.

Reel · Capabilities

WAN 3.0's Powerful Features

What Alibaba's unified video model brings to text, image and reference driven generation

  1. Feature 01 / 08

    One Unified Model

    WAN 2.7 split text-to-video, image-to-video, reference generation and editing across separate models. WAN 3.0 folds them into a single model, so behaviour stays consistent whichever mode you start from.

  2. Feature 02 / 08

    30 Seconds in One Pass

    Set any duration from 2 to 30 seconds and the whole clip renders in a single generation, making continuous camera moves and one-take shot language possible without stitching separate clips together.

  3. Feature 03 / 08

    Native Synchronized Audio

    Audio is generated with the picture rather than added afterwards, covering dialogue timbre, environmental sound and musical rhythm, so speech and action line up without a separate pass.

  4. Feature 04 / 08

    480p to 1080p Output

    Pick 480p for quick drafts, 720p for everyday delivery or 1080p for final work. Billing is per second at each tier, so short tests stay cheap and only finals cost full price.

  5. Feature 05 / 08

    Omni-Reference Inputs

    Reference-to-Video accepts up to 10 reference images, 5 reference video clips and 5 audio tracks in one run, letting you lock a character, a product or a location across a whole sequence.

  6. Feature 06 / 08

    First and Last Frame Control

    Image-to-Video animates the first-frame image you upload and optionally accepts a last-frame image, so you can pin exactly where a shot begins and where it should land on screen.

  7. Feature 07 / 08

    Deep-Thinking Mode

    Turn on thinking mode and the model reasons more deliberately about a complex prompt before it starts generating, which helps with multi-beat scenes and unusual staging instructions.

  8. Feature 08 / 08

    Five Aspect Ratios

    Choose 16:9, 9:16, 1:1, 4:3 or 3:4 to cover widescreen delivery, vertical mobile, square social feeds and classic framing from one parameter. Image-to-Video can also follow the input image.

How to Use WAN 3.0 Text to Video

Generate a clip of up to 30 seconds from a written description

Write Your Prompt

Describe the scene, subject, camera movement, lighting and mood. Detail helps, and you can turn on thinking mode for complex ideas.

FAQ

Frequently Asked Questions

Common questions about WAN 3.0 video generation

WAN 3.0 is Alibaba Tongyi Lab's unified video generation model, served on Dreamega through WaveSpeed. WAN 2.7 split text-to-video, image-to-video, reference generation and editing across separate models; WAN 3.0 folds those capabilities into one model. The practical differences are single-pass clips of up to 30 seconds, native synchronized audio, and an optional deep-thinking mode for complex prompts.
Three. Text-to-Video builds a clip from a written prompt. Image-to-Video animates a first-frame image you upload and optionally accepts a last-frame image. Reference-to-Video combines your prompt with up to 10 reference images, 5 reference videos and 5 reference audio tracks to keep a subject consistent across the shot.
Output is 480p, 720p or 1080p. WAN 3.0 is not a 4K model. Duration is any whole number of seconds from 2 to 30, and the entire clip renders in a single generation pass rather than being stitched from shorter segments.
Yes. Audio generation is on by default and is produced together with the picture, covering dialogue timbre, environmental sound and musical rhythm so speech and action stay in sync. You can switch it off if you only want silent footage to score yourself.
Reference-to-Video takes up to 10 reference images, up to 5 reference videos and up to 5 reference audio tracks, and at least one of the three is required. Reference video and audio are each capped at 15 seconds in total. Refer to items in your prompt by their order, for example "the woman in image 1".
Billing is per second of output. Text-to-Video and Reference-to-Video cost 11 credits per second at 480p, 20 at 720p and 42 at 1080p, so a 5-second 720p clip is 100 credits. Image-to-Video is cheaper at 9, 18 and 36 credits per second for the same tiers.
Pricing · Choose Yours

Flexible AI Pricing

Pay-as-you-go credits or subscription plans. No hidden fees, cancel anytime.

One Time supports crypto payment (BTC, USDT, ETH, 350+)

Monthly billing

Free

Try before you buy

0
One Time
USD
Free
32points
Up to 3 videos
Up to 32 images
Multi-Model Support
Text to Video
Image to Video
Video to Video
Consistent Character
AI Animation Generator
Templates & Effects
AI Video Enhancers
Interactive Community
Faster Generation Speed
No-watermark Outputs
More Camera Movement
Private Video Visibility
Copy Protection
Priority Support
Popular

Pro

Elevate your AI experience

29.99
1 Month
USD
800
800points1 Month
Up to 80 videos1 Month
Up to 800 images1 Month
3 tasks(Parallel Tasks)
Multi-Model Support
Text to Video
Image to Video
Video to Video
Consistent Character
AI Animation Generator
Templates & Effects
AI Video Enhancers
Interactive Community
Faster Generation Speed
No-watermark Outputs
More Camera Movement
Private Video Visibility
Copy Protection
Priority Support

Lite

Start your AI journey

19.99
1 Month
USD
300points1 Month
Up to 30 videos1 Month
Up to 300 images1 Month
3 tasks(Parallel Tasks)
Multi-Model Support
Text to Video
Image to Video
Video to Video
Consistent Character
AI Animation Generator
Templates & Effects
AI Video Enhancers
Interactive Community
Faster Generation Speed
No-watermark Outputs
More Camera Movement
Private Video Visibility
Copy Protection
Priority Support