Wan 2.5 Preview is here — explore its top 6 big new features, including audio-visual sync, 1080p video, conversational image editing, and its release date.

By Wan AI Team
24 Sep 2025
12
423K
Wan 2.5 Release Date
September 24, 2025 (Preview)
Wan 2.5 Preview became available in September 2025. A separate final-release date has not been confirmed, so this page will be updated when an official date is announced.
Alibaba has officially launched its next-generation AI Model, WAN 2.5 Preview. This release marks a significant step forward for AI in video and image generation, with its new architecture and powerful features set to revolutionize how we create and edit visual content.
Imagine crafting a cinematic short film with perfect music, sound effects, and character dialogue, all with just a few simple commands. This is no longer science fiction; it's the reality brought to you by WAN 2.5 Preview. Best of all, you don't need any special equipment or complex setups. The Model offers a fast online experience that brings your creative ideas to life instantly.
Here is a quick look at the biggest new features that make Wan 2.5 Preview a major step up over previous versions:
For creators searching for Wan 2.5 image capabilities, the biggest change is that Wan 2.5 brings image generation and conversational image editing into the same multimodal creative workflow while also improving video quality and duration.
| Core feature | Wan 2.1 | Wan 2.5 Preview |
|---|---|---|
| Image generation | Text-to-image support in the broader model suite | Enhanced text-to-image generation with stronger prompt following |
| Image editing | Video editing and image-conditioned workflows | Conversational image editing, including object, color, and background changes |
| Text-to-video | Supported | Supported with improved motion and structural stability |
| Image-to-video | Supported on dedicated I2V models | Supported with improved image understanding and visual consistency |
| Maximum video length | Typically 5 seconds | Up to 10 seconds |
| Maximum video resolution | Up to 720p on 14B models; 480p on the 1.3B model | Up to 1080p |
| Native audio generation | Not a core generation feature | Synchronized dialogue, music, ambience, and sound effects |
| Audio input | Not a standard input mode | Audio-conditioned video generation |
| Availability | Open-source model weights and local deployment | Online preview |
In short, Wan 2.1 remains a strong choice for open-source and local workflows, while Wan 2.5 Preview is aimed at creators who want higher-resolution output, longer clips, native sound, and more capable image creation and editing.
The most remarkable innovation in WAN 2.5 Preview is its ability to overcome the common "out-of-sync" problem in AI-generated video, achieving perfect audio-visual synchronization.
How to Experience: All these powerful video features are available on the WAN 2.5 Preview online Model. Just visit the website, enter text or upload audio, and effortlessly generate your own cinematic videos without worrying about technical configurations.
Beyond its impressive video capabilities, WAN 2.5 Preview also offers a huge leap in image creation.
How to Experience: Log in to the Model now, upload an image, and try editing it with simple, conversational commands. You'll be amazed by the AI's understanding and its ability to deliver precise, pixel-perfect results.
A big leap forward for image & video generation
These groundbreaking features are all thanks to WAN 2.5 Preview's new unified multimodal architecture. This is like giving the AI a single brain that can simultaneously see, hear, read, and write, integrating the processing of text, images, video, and audio into one seamless framework. Through joint multimodal training, the model achieves stronger alignment between different data types, which is crucial for perfect audio-visual sync and precise instruction-following.
In short, WAN 2.5 Preview is more than just a tool; it's an intelligent creative partner. It makes professional-level video and image creation incredibly simple, putting the power of creativity back into the hands of everyone.
Wan 2.5 Preview became available in September 2025. It is currently in preview mode, and Alibaba has not confirmed a separate final-release date yet. We will update this article as soon as an official date is announced.
The biggest new features are native audio-visual synchronization, 10-second 1080p video generation, audio conditioning as an input, and conversational image editing. Together, they let you create cinematic videos with dialogue, music, and sound effects directly in your browser.
Wan 2.5 Preview adds native audio generation, audio-input video generation, conversational image editing, longer clips (up to 10 seconds), and up to 1080p resolution. Wan 2.1 remains a strong choice for open-source, local, and offline workflows.
Yes. Wan 2.5 Preview can generate synchronized dialogue, music, ambience, and sound effects alongside your video. You can also provide audio as an input, and the model will generate a video that matches its mood and atmosphere.
Yes. Wan 2.5 Preview supports conversational image editing. You can give simple instructions — such as "change the car's color to blue" or "replace the background with a snowy mountain" — and the AI will apply precise, pixel-level edits.
Yes. Wan 2.5 Preview supports image-to-video generation with improved image understanding and visual consistency compared with Wan 2.1, so your generated clips stay faithful to the reference image.
Wan 2.5 Preview is available through an online preview that requires no installation. Visit the Wan AI website, enter your text prompt or upload an audio/image, and start generating — no technical setup needed.
No. Unlike Wan 2.1's local deployment option, Wan 2.5 Preview runs online in the browser, so you don't need a high-end GPU or any special equipment to use its features.
Wan 2.5 Preview supports videos up to 10 seconds long and up to 1080p resolution with cinematic quality and stable motion.