Try these ideas:
🎬 One video · $2.99 · no sign-up — we’ll email you the link. Already have credits? Sign in
WAN 2.7 by Alibaba brings 7 generation modes in one model — text-to-video, image-to-video, start-end animation, video continuation, AI video editing, audio-driven video, and multi-reference consistency. Generate 720p or 1080p videos up to 15 seconds with native audio on VO3 AI.
Try WAN 2.7 text-to-video and image-to-video directly. Select WAN 2.7 from the model dropdown to experience Alibaba's latest AI video technology.
Generate AI Model videos with AI
Try these ideas:
🎬 One video · $2.99 · no sign-up — we’ll email you the link. Already have credits? Sign in
WAN 2.7 is Alibaba's most advanced AI video generation model, representing a major leap from WAN 2.5. While WAN 2.5 offered basic text-to-video and image-to-video with fixed 5-second output, WAN 2.7 introduces 7 distinct generation modes with configurable duration (10-15 seconds), dual resolution options (720p/1080p), and native audio generation. Built on Alibaba's latest research, WAN 2.7 excels at maintaining visual consistency across complex multi-reference scenarios.
WAN 2.7 is available exclusively on VO3 AI through the KIE API infrastructure. Unlike single-purpose models, WAN 2.7 serves as a complete video creation toolkit — from generating new content to editing existing videos, continuing scenes, and even driving visuals with audio input. The Reference-to-Video (R2V) mode uniquely supports up to 5 combined image and video references for unprecedented character and style consistency.
Generate 720p/1080p videos up to 15s
from text prompts with native audio
Animate single images with first-frame
control and aspect ratio options
Upload start and end frames —
WAN 2.7 generates smooth motion between them
Extend existing video clips
seamlessly while maintaining style consistency
Edit videos with natural language instructions
— change style, background, or mood
Drive video generation with audio files
— sync visuals to beats and rhythm
Up to 5 image/video references +
voice for character and style consistency
Supports both English and Chinese
prompts with intelligent prompt extension
TOOLS & DEMOS
Each mode shown with its input, prompt, and output. Click any card to try it yourself.
Dual resolution options for
both speed-optimized and quality-focused workflows
Configurable video length with
duration-based pricing for cost control
Built-in audio synthesis that
matches generated visuals automatically
Intelligent prompt rewriting that expands
brief descriptions into detailed scenes
Optimized pipeline delivers results in
2-5 minutes depending on settings
R2V mode maintains character
identity across multiple reference materials
Credits start from $2.99 · View all plans
7 generation modes, 720p/1080p output, native audio. The most versatile AI video model available on VO3 AI.
WAN 2.7 AI video generator | Alibaba WAN 2.7 model | WAN 2.7 text to video | WAN 2.7 image to video | WAN 2.7 video continue | WAN 2.7 video edit | WAN 2.7 audio to video | WAN 2.7 reference to video | AI video generator 2026 | VO3 AI WAN 2.7