The Wan 2.7 video model enables one‑stop unified generation of text‑driven, image‑driven, and reference‑driven video content along with native synchronized audio. It delivers sharper details, smoother cinematic motion effects, and more coherent shot language, allowing large‑scale production of narrative content suitable for professional creation.
WAN 2.7 video editing enables prompt-based video editing, supports multi-image reference, and can output videos with 720p/1080p resolution.
WAN 2.7 Image to Video supports multimodal input (text/image/audio/video) and can perform three major tasks: video generation from the first frame, video generation from the first and last frames, and video continuation.
The WAN 2.7 text-to-video model can transform simple prompts into coherent, cinematic video clips with clear details and stable camera movement. It strictly follows instructions and is suitable for advertising, explanatory videos, and social media content creation.