wan 2.2
wan 2.2, image-to-video, 5.0s-480pVACE Fun for Wan 2.2 A14B from Alibaba-PAI
fal-ai/wan-22-vace-fun-a14b/outpainting
Source checked: 2026-09-13 · Provider endpoint, not an independently benchmarked model.
These are source observations, not interchangeable quotes. Token, compute-time, output-duration and per-request prices cannot be compared by the number alone. Input charges and optional features may add to the total.
Your request will cost $0.10 per video second for 720p, $0.075 per video second for 580p, $0.05 per video second for 480p. Video seconds are calculated at 16 frames per second.
| category | video-to-video |
|---|---|
| license | commercial |
| first_frame_url | URL to the first frame of the video. If provided, the model will use this frame as a reference. |
| video_quality | The quality of the generated video. · Options: low, medium, high, maximum · Default: high |
| auto_downsample_min_fps | The minimum frames per second to downsample the video to. This is used to help determine the auto downsample factor to try and find the lowest detail-preserving downsample factor. The default value is appropriate for most videos, if you are using a video with very fast motion, you may need to increase this value. If your video has a very low amount of motion, you could decrease this value to allow for higher downsampling and thus longer sequences. · Min: 1 · Max: 60 · Default: 15 |
| num_frames | Number of frames to generate. Must be between 81 to 241 (inclusive). · Min: 17 · Max: 241 · Default: 81 |
| frames_per_second | Frames per second of the generated video. Must be between 5 to 30. Ignored if match_input_frames_per_second is true. · Default: 16 |
| match_input_frames_per_second | If true, the frames per second of the generated video will match the input video. If false, the frames per second will be determined by the frames_per_second parameter. · Default: false |
| num_interpolated_frames | Number of frames to interpolate between the original frames. A value of 0 means no interpolation. · Min: 0 · Max: 5 · Default: 0 |
| resolution | Resolution of the generated video. · Options: auto, 240p, 360p, 480p, 580p, 720p · Default: auto |
| return_frames_zip | If true, also return a ZIP file containing all generated frames. · Default: false |
| last_frame_url | URL to the last frame of the video. If provided, the model will use this frame as a reference. |
| match_input_num_frames | If true, the number of frames in the generated video will match the number of frames in the input video. If false, the number of frames will be determined by the num_frames parameter. · Default: false |
| aspect_ratio | Aspect ratio of the generated video. · Options: auto, 16:9, 1:1, 9:16 · Default: auto |
Unlisted settings, region availability, deposits, taxes and commercial terms remain unverified. Confirm the exact endpoint before purchase.
Fast, Pro, editing and audio variants remain separate. Match the exact configuration before comparing costs.
Source-listed endpoints and pricing conditions. Variants are not unique base models; account quotes do not enter cheapest-price rankings.
Counts describe collected endpoints, not unique models or a guarantee of complete coverage. Price evidence may be a numeric rate or source billing text. Providers with no published endpoints are still awaiting collection.
58 endpoints · 1/3
Variants listed separately. Open a model to check units and conditions.
wan 2.2, image-to-video, 5.0s-480pwan 2.2, image-to-video, 5.0s-480p
wan 2.2, image-to-video, 5.0s-720pwan 2.2, image-to-video, 5.0s-720p
wan 2.2, image-to-video, 5.0s-580pwan 2.2, image-to-video, 5.0s-580p
wan 2.2, text-to-video, 5.0s-580pwan 2.2, text-to-video, 5.0s-580p
wan 2.2, text-to-video, 5.0s-480pwan 2.2, text-to-video, 5.0s-480p
wan 2.2, text-to-video, 5.0s-720pwan 2.2, text-to-video, 5.0s-720p
Wan 2.2 A14B Turbo API Speech to Video, 480pWan 2.2 A14B Turbo API Speech to Video, 480p
Wan 2.2 A14B Turbo API Speech to Video, 720pWan 2.2 A14B Turbo API Speech to Video, 720p
Wan 2.2 A14B Turbo API Speech to Video, 580pWan 2.2 A14B Turbo API Speech to Video, 580p
wan 2.2 Animate, 2.2 Animate Replace, 1.0s-720pwan 2.2 Animate, 2.2 Animate Replace, 1.0s-720p
wan 2.2 Animate, 2.2 Animate Replace, 1.0s-580pwan 2.2 Animate, 2.2 Animate Replace, 1.0s-580p
wan 2.2 Animate, 2.2 Animate Replace, 1.0s-480pwan 2.2 Animate, 2.2 Animate Replace, 1.0s-480p
wan 2.2 Animate, 2.2 Animate Move, 1.0s-480pwan 2.2 Animate, 2.2 Animate Move, 1.0s-480p
wan 2.2 Animate, 2.2 Animate Move, 1.0s-580pwan 2.2 Animate, 2.2 Animate Move, 1.0s-580p
wan 2.2 Animate, 2.2 Animate Move, 1.0s-720pwan 2.2 Animate, 2.2 Animate Move, 1.0s-720p
fal-ai/wan-22-vace-fun-a14b/depthVACE Fun for Wan 2.2 A14B from Alibaba-PAI
fal-ai/wan-22-vace-fun-a14b/inpaintingVACE Fun for Wan 2.2 A14B from Alibaba-PAI
fal-ai/wan-22-vace-fun-a14b/outpaintingVACE Fun for Wan 2.2 A14B from Alibaba-PAI
fal-ai/wan-22-vace-fun-a14b/reframeVACE Fun for Wan 2.2 A14B from Alibaba-PAI
fal-ai/wan/v2.2-14b/animate/moveWan-Animate is a video model that generates high-fidelity character videos by replicating the expressions and movements of characters from reference videos.
fal-ai/wan/v2.2-14b/animate/replaceWan-Animate Replace is a model that can integrate animated characters into reference videos, replacing the original character while preserving the scene’s lighting and color tone for seamless environm
fal-ai/wan/v2.2-14b/speech-to-videoWan-S2V is a video model that generates high-quality videos from static images and audio, with realistic facial expressions, body movements, and professional camera work for film and television applic
Model sample · hivamohfal-ai/wan/v2.2-a14b/text-to-videoWan-2.2 text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.
fal-ai/wan/v2.2-a14b/text-to-video/loraWan-2.2 text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts. This endpoint supports LoRAs made for Wan 2.2.