Wan 2.6
Wan 2.6 is a video generation model that creates videos from source images and audio with multi-shot control.
Image and audio-driven video generation
Wan 2.6 is a video generation model published by Wan and served through the WaveSpeed provider on MindStudio. It accepts image and audio inputs alongside text prompts to produce video output, and supports configurable resolution, duration, and shot type selection. The model was released in December 2025 and is priced at $0.10–$0.15 per second of generated video.
Wan 2.6 is designed for workflows that require animating a source image with accompanying audio, making it suited for tasks like lip-sync video, scene animation, and multi-shot video production. Its shot type toggle lets users specify cinematic framing, and a seed input enables reproducible outputs. The 1000-token context window applies to the text prompt inputs used to guide generation.
What Wan 2.6 supports
Image-to-Video
Animates a source image into video using a provided image URL as the visual starting point for generation.
Audio-Driven Generation
Accepts an audio input to synchronize or influence the generated video, enabling use cases like lip-sync and audio-reactive animation.
Multi-Shot Control
Supports a shot type toggle that lets users specify different cinematic shot framings within a single generation request.
Resolution Selection
Offers selectable output resolution options so users can match video quality to their delivery requirements.
Duration Control
Allows users to select the length of the generated video clip, with pricing applied per second of output at $0.10–$0.15.
Negative Prompting
Accepts a negative prompt text input to suppress unwanted visual elements or styles from the generated video.
Reproducible Outputs
Includes a seed input that allows users to reproduce identical generation results by reusing the same seed value.
Ready to build with Wan 2.6?
Get Started FreeCommon questions about Wan 2.6
How is Wan 2.6 priced?
Wan 2.6 is priced at $0.10 to $0.15 per second of generated video. The exact rate within that range may depend on resolution or other generation settings.
What inputs does Wan 2.6 accept?
Wan 2.6 accepts a source image URL, an audio input, a text prompt, a negative prompt, resolution selection, duration selection, shot type selection, and a seed value.
What is the context window for Wan 2.6?
Wan 2.6 has a context window of 1000 tokens, which applies to the text prompt inputs used to guide video generation.
What is the knowledge cutoff or release date for Wan 2.6?
Wan 2.6 was released in December 2025. No specific training data cutoff date is listed in the available metadata.
Can I control the shot framing in generated videos?
Yes. Wan 2.6 includes a shot type toggle input that lets you specify different cinematic shot types, supporting multi-shot video production workflows.
Does Wan 2.6 support reproducible video generation?
Yes. Wan 2.6 includes a seed input. Using the same seed with the same inputs will reproduce the same generated video output.
What people think about Wan 2.6
Community discussion around Wan 2.6 on Reddit was generally positive, with users highlighting its native audio synchronization and 1080p output as notable features for a single-pass video generation model. The thread, which received 226 upvotes and 78 comments, noted that the model appeared on API platforms ahead of Alibaba's official announcement event.
Some commenters framed the release in the context of the broader competitive landscape for video generation models, while others focused on practical use cases such as short-form content and automated video production. No widespread technical limitations were documented in the available thread data.
Parameters & options
Description of what to exclude from the video.
A specific value that is used to guide the 'randomness' of the generation.
Explore similar models
Start building with Wan 2.6
No API keys required. Create AI-powered workflows with Wan 2.6 in minutes — free.