Alibaba: Wan 3.0
alibaba/wan-3.0Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
alibaba/wan-3.0/text-to-video
Source checked: 2026-09-13 · Provider endpoint, not an independently benchmarked model.
These are source observations, not interchangeable quotes. Token, compute-time, output-duration and per-request prices cannot be compared by the number alone. Input charges and optional features may add to the total.
For every second of video you generate, you will be charged $0.05 480p, $0.10 720p, or $0.20 1080p. For example, a 5s video at 720p will cost $0.50. Pricing is subjected to change.*
| category | text-to-video |
|---|---|
| license | commercial |
| duration | Output duration in seconds. Set to null for smart duration, which lets the model pick a length from the prompt and reference media. · Default: 5 |
| resolution | Output video resolution tier. · Options: 480p, 720p, 1080p · Default: 1080p |
| aspect_ratio | Output aspect ratio, or adaptive selection. · Options: adaptive, 16:9, 4:3, 1:1, 3:4, 9:16 · Default: adaptive |
| audio | Include generated audio. · Default: true |
Unlisted settings, region availability, deposits, taxes and commercial terms remain unverified. Confirm the exact endpoint before purchase.
Fast, Pro, editing and audio variants remain separate. Match the exact configuration before comparing costs.
Source-listed endpoints and pricing conditions. Variants are not unique base models; account quotes do not enter cheapest-price rankings.
Counts describe collected endpoints, not unique models or a guarantee of complete coverage. Price evidence may be a numeric rate or source billing text. Providers with no published endpoints are still awaiting collection.
23 endpoints · 1/1
Variants listed separately. Open a model to check units and conditions.
alibaba/wan-3.0Wan 3.0 is a video generation model from Alibaba for text-to-video, image-to-video, and reference-guided video generation. It produces 480p, 720p, or 1080p video with durations from 2 to 30 seconds.
alibaba/wan-3.0-primeWan 3.0 Prime is a fast-mode variant of [Wan 3.0](https://openrouter.ai/alibaba/wan-3.0) from Alibaba. It supports text-to-video and first-frame image-to-video generation.
alibaba/wan-3.0-prime/image-to-videoWan 3.0 Prime Image to Video is an accelerated variant that animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control,
alibaba/wan-3.0-prime/reference-to-videoWan 3.0 Prime Reference to Video is an accelerated variant that creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration
alibaba/wan-3.0-prime/text-to-videoWan 3.0 Prime Text to Video is an accelerated variant that generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking con
alibaba/wan-3.0/image-to-videoWan 3.0 Image to Video animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking co
alibaba/wan-3.0/reference-to-videoWan 3.0 Reference to Video creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration and aspect ratio control for subject
alibaba/wan-3.0/text-to-videoWan 3.0 Text to Video generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generati
alibaba/wan-3.0/reference-to-videoWan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
alibaba/wan-3.0/image-to-videoWan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
alibaba/wan-3.0-prime/reference-to-videoWan 3.0 Prime Reference-to-Video combines reference images, videos, and audio into a unified video with fast generation and strong multimodal coherence. It follows character identity, visual style, mo
alibaba/wan-3.0-prime/image-to-videoWan 3.0 Prime Image-to-Video turns still images into dynamic, cinematic sequences with rapid turnaround, natural motion, and excellent visual continuity. It preserves the identity, composition, and at
alibaba/wan-3.0-prime/text-to-videoWan 3.0 Prime Text-to-Video transforms written prompts into polished videos with accelerated generation, fluid motion, strong scene fidelity, and coherent visual storytelling. Built for fast creative
wan 3.0 video, 1080p, videowan 3.0 video, 1080p, video
wan 3.0 video, 720p, videowan 3.0 video, 720p, video
wan 3.0 video, 480p, videowan 3.0 video, 480p, video
alibaba/wan-3.0/text-to-videoWan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
alibaba:wan@3.0Wan3.0
alibaba:wan@3.0-primeWan3.0 Prime
wan3.0 video prime, 1080p, videowan3.0 video prime, 1080p, video
wan3.0 video prime, 720p, videowan3.0 video prime, 720p, video
wan3.0 video prime, 480p, videowan3.0 video prime, 480p, video
wan3.0-videowan3.0-video · per_second