Alibaba
WAN 3.0 AI video generator
Alibaba's newest video model. Up to 30 seconds with native audio, from 480p to 1080p.
Free credits on signup, no card required. Plans from $19 one time.

Example videos made on Enhance AI
These videos come from related models on the same account. The model that made each one is named under it.
A lone kayaker paddling across a glassy alpine lake at dawn, mist lifting off the water, pine forest mirrored in the surface. Static wide shot on a tripod, moody low key lighting, fine film grain. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
A red hot air balloon drifting over patchwork farmland, long morning shadows, a flock of birds crossing below. Slow dolly push in, misty dawn light, desaturated greens. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
Waves breaking against black volcanic rocks, spray hanging in the air, a lighthouse on the headland. Slow pull back revealing the wider scene, overcast soft light, muted natural palette. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
A cyclist descending a winding mountain pass, switchbacks below, clouds moving through the valley. Slow orbit around the subject, soft window light, gentle contrast. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
A herd of wild horses galloping across a dusty plain, manes flying, dust glowing in low sun. Aerial view gliding forward, golden hour light, warm colour grade. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
An old steam train crossing a stone viaduct through autumn woodland, smoke trailing behind. Crane shot rising from ground level, harsh midday sun, deep shadows. Cinematic and photorealistic, natural motion. Photographic, natural, no text, no lettering, no captions, no subtitles, no watermark, no logo, no brand names, no product labels, no user interface, no screens, no recognisable fictional characters, no celebrities or recognisable real people.
Generated with MiniMax H3 Max
Key features
WAN 3.0 turns a written prompt and a still image you upload into video inside Enhance AI. It produces clips from 5 to 30 seconds at a resolution you choose, 720p by default. On top of that it renders 16:9, 9:16, 1:1, 4:3, 3:4 and adaptive. WAN 3.0 is built by Alibaba. On Enhance AI it runs on the same account as every other model, with no separate subscription.
Clip length
WAN 3.0 renders clips from 5 to 30 seconds, and you pick the length before you render. Longer clips cost more because video is billed by length, so rough out an idea short and only render the full length once the shot works.
Resolution
Choose 480p, 720p or 1080p before you render. A lower resolution costs less, so test a prompt low and rerender the keeper at the top setting. The studio shows the price for the setting you picked before anything is charged.
Input modes
Start from a written prompt or from a still image. Text to video gives WAN 3.0 the most freedom; image to video locks the first frame, so a product, a face or a location stays as you shot it while the motion is generated around it.
Aspect ratios
WAN 3.0 renders 16:9, 9:16, 1:1, 4:3, 3:4 and adaptive. Pick the ratio for the platform before you render rather than cropping afterwards: 9:16 for Reels and Shorts, 16:9 for YouTube and 1:1 for feeds. The whole frame is generated for that shape.
Audio
WAN 3.0 can generate sound with the picture, so the clip arrives with audio that matches what is on screen rather than silence you have to fill later. You can still mute it and lay your own voice or music track over the top in any editor.
Specifications
- Provider
- Alibaba
- Type
- Video
- Released
- 24 August 2026
- Resolution
- 720p by default
- Modes
- Text to video, Image to video
- Duration
- 5 to 30 seconds
- Aspect ratio
- 16:9, 9:16, 1:1, 4:3, 3:4, adaptive
How to use WAN 3.0 on Enhance AI
- 1
Sign in and pick WAN 3.0
Open the Video Studio on Enhance AI and choose WAN 3.0 from the model list. A new account starts with free credits and no card is needed. The Try button on this page opens the studio with WAN 3.0 already selected.
- 2
Describe the clip or upload a frame
Write a prompt that says what is in the shot, where it is and how the camera moves, or upload a still image to animate as the first frame. Then pick a length from 5 to 30 seconds, a resolution and an aspect ratio.
- 3
Generate, then download or extend
Press Generate. WAN 3.0 renders in the cloud and the clip lands in your history on Enhance AI, where you can download the MP4, upscale it or extend it with another model on the same account. You see the price before you confirm.
What people make with WAN 3.0
Backgrounds and loops
Atmospheric loops for a website header, a stage screen or a stream overlay suit WAN 3.0 well: slow drifting clouds, rain on glass, light moving across a wall. Keep the camera still and the motion gentle so the loop point hides, and render at the ratio the screen needs.
Short social clips
Vertical clips for Reels, TikTok and Shorts are the most common job for WAN 3.0. Write one clear scene per clip, keep the action simple so the motion reads on a phone, and render a few variations before you post. A short clip that lands is worth more than a long one that wanders.
Product shots
Give WAN 3.0 a product, the surface it sits on and one camera move, and it returns a short hero shot for a store page or an ad. Keep the prompt to what the camera sees. Start from a product photo so the packaging stays exact.
Common questions
- What is WAN 3.0?
- WAN 3.0 is a video generation model from Alibaba. It turns a written prompt or a still image into video: 30 second clips at 1080p with native audio. On Enhance AI it runs on the same account as 250+ other models, with no separate subscription.
- Can WAN 3.0 turn a photo into a video?
- Yes. Upload a still image, describe the movement you want, and WAN 3.0 animates it. It also works from a written prompt alone.
- When was WAN 3.0 released?
- Alibaba released WAN 3.0 on 24 August 2026. You can run it on Enhance AI now, along with every other video model on one account.
- What is new in WAN 3.0 compared to WAN 2.7?
- On Enhance AI, WAN 2.7 renders clips up to 15 seconds. WAN 3.0 renders clips up to 30 seconds and adds native audio. Both run on the same account, so you can pick whichever fits the shot and the budget.
- How long can a WAN 3.0 clip be?
- WAN 3.0 produces clips from 5 to 30 seconds. Longer runs cost more, because video is billed by length rather than by generation.
- What resolution does WAN 3.0 output?
- WAN 3.0 renders at a resolution you choose, 720p by default. You pick the resolution before you render, and a lower one costs less.
- Does WAN 3.0 generate audio?
- Yes. WAN 3.0 can generate sound with the picture, so the clip arrives with audio rather than silence. Mute it and lay your own track over the top if you prefer.
- Is WAN 3.0 free to try?
- Yes. Free credits on signup, no card required, so you can run WAN 3.0 before paying anything. Plans from $19 one time when you need more, and one account covers WAN 3.0 and 250+ other models.
- What does it cost to run WAN 3.0?
- Video renders are billed from your wallet by clip length and resolution, so a short clip at a lower resolution costs less than a long one at full quality. You see the price before you render.
- Do I need a separate subscription for WAN 3.0?
- No. WAN 3.0 runs on the same Enhance AI account as every other model, so you are not paying one provider for video and another for images.
More from Alibaba
- WAN 3.0 PrimeThe higher quality tier of WAN 3.0. Same controls, sharper results: up to 30 seconds with native audio, from 480p to 1080p.View
- WAN 3.0 VibeAlibaba's newest video model with a deep thinking mode for complex prompts. Up to 30 seconds with native audio, from 480p to 1080p.View
- WAN 3.0 Vibe ProThe higher quality tier of WAN 3.0 Vibe, with a deep thinking mode for complex prompts. Up to 30 seconds with native audio, from 480p to 1080p.View
- WAN 2.7Alibaba's WAN 2.7 model. High quality 720p or 1080p Image to Video with customizable aspect ratios.View
- HappyHorse 1.1Alibaba HappyHorse 1.1. Cinematic 720p or 1080p video from text or image with smooth motion.View
- WAN 2.6Alibaba's WAN 2.6 model with flexible resolution and duration optionsView
Other video generation models
More from Enhance AI