Alibaba logo

Wan 2.7

AlibabaReleased Apr 1, 2026wanx/wan-v2-7
Compare

Wan 2.7 is Alibaba Tongyi Lab's flagship video generation model, offering text-to-video, image-to-video, reference-to-video, and instruction-based video editing through one endpoint. It supports first-and-last-frame control, up to 5 reference images, videos, and audio clips for consistent subject and voice identity, native synchronized audio (ambient sound, lip-synced dialogue, music), connected multi-shot generation, and clips of 2-15 seconds at up to 1080P.

Context
4K
Max output
4K
Input /1M
—
Output /1M
$0.1/vid
AcceptsTextImageVideoAudioProducesVideoInstruct type none

Providers

FastRouter routes your requests to this provider. You can also bring your own key.

Alibaba logo
Alibaba
alibaba
Input /1M—
Output /1M$0.1/vid
Context4K
Max output4K

Supported parameters

Request parameters you can send with this model.

Parameter
Type
Alibaba
Core (Sampling & basic generation)
seedSeed for deterministic sampling.
number
0 to 2147483647
Media (Generation Controls)
aspectRatioOutput image aspect ratio.
string
"16:9", "9:16", "1:1", "4:3", "3:4"
Other
audioUrl
string
-
duration
number
2 to 15
enhancePrompt
bool
true | false
first_frame
string
-
generateAudio
bool
true | false
image
string
-
images
array
max 5 items
last_frame
string
-

Make your first API call

OpenAI-compatible. Point your SDK at FastRouter, or call the API directly.

Full API docs
Replace <FASTROUTER_API_KEY> with your key · Get your key →
HTTP

Frequently asked questions

More models from Alibaba

Alibaba logo
HappyHorse 1.1alibaba/happyhorse-1.1

HappyHorse 1.1 is Alibaba's closed-weight cinematic video generation model, producing clips of 3-15 seconds with synchronized audio and multilingual lip-synced dialogue generated in a single pass. It generates video from a text prompt, a single starting image, or a set of reference images, and supports video editing of existing clips, at output resolutions of 480p, 720p, and 1080p. It improves on HappyHorse 1.0 with stronger prompt adherence, smoother motion, and more consistent characters across frames, and previously topped the Artificial Analysis video arena anonymously before Alibaba was revealed as its maker.

Alibaba logo
Qwen3.8 27Bqwen/qwen3.8-27b

Qwen3.8-27B is Alibaba's compact, deployment-friendly open-weight model from the Qwen3.8 generation, released under Apache 2.0. It is a 27.8-billion-parameter dense model using a hybrid decoder that alternates Gated DeltaNet linear attention with grouped-query full attention, with native text, image, and video understanding, a 262,144-token context window, and configurable reasoning effort toggled on or off per request. It targets coding, professional work, research, and long-horizon agentic tasks in a footprint small enough to self-host on high-end consumer hardware.

Alibaba logo
Qwen3.8 Maxqwen/qwen3.8-max

Qwen3.8-Max is Alibaba's flagship sparse mixture-of-experts model, with 2.4 trillion total parameters and approximately 95 billion active per forward pass, accepting text, image, and video input and returning text with a 1M-token context window. Hybrid thinking mode is enabled by default with three reasoning-effort tiers, and pricing is flat across the full context window with no long-prompt surcharge. Target workloads include software engineering, long-horizon agentic work, office productivity, and visual tasks such as turning screenshots or design files into working pages.

Alibaba logo
Wan 2.6 wanx/wan-v2-6

Wan 2.6 is a state-of-the-art, open-source multimodal AI video generation model developed by Alibaba Cloud (released around December 2025). It is designed to create high-fidelity, 1080p cinematic videos up to 15 seconds long from text, images, or reference videos.

Model scores sourced from ArtificialAnalysis