Vidu Image
Image editing
Vidu Reference-to-Image creates images by using a reference images and combining them with a prompt.
Find a model. See what it can do.
Browse every available model. Search by name or task.
Image editing
Vidu Reference-to-Image creates images by using a reference images and combining them with a prompt.
Text to video · Image to video
Vidu Q1 Text to Video generates high-quality 1080p videos with exceptional visual quality and motion diversity
Text to image
Wan 2.2's 5B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail
Text to video
Wan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
Text to image · Image editing
Wan 2.2's 14B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail
Text to video · Image to video
Wan-2.2 turbo text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.
Text to image · Image editing
Generate premium-quality images from text prompts using the enhanced WAN 2.7 Pro model with superior detail and composition.
Text to image · Image editing
Generate high-quality images from text prompts using the WAN 2.7 model with advanced prompt understanding and detailed output.
Text to image · Image editing
Wan 2.5 text-to-image model.
Text to video · Image to video
Wan 2.5 text-to-video model.
VideoText to video
Wan 2.6 Uncensored for video generation with optional input media.
Text to image
Wan 2.7 Image for image generation and editing.
VideoImage to video
Animate one input image into a 5, 10 or 15 second video.
Image to video
Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
VideoText to video
Wan 3.0 for video generation with optional input media.
Image to video
Wan 3.0 Prime Image-to-Video turns still images into dynamic, cinematic sequences with rapid turnaround, natural motion, and excellent visual continuity. It preserves the identity, composition, and atmosphere of the source image while introducing expressive movement, camera dynamics, and richly detailed animation.
VideoText to video
Wan 3.0 Prime for video generation with optional input media.
VideoText to video
Wan 3.0 Prime Pro Uncensored for video generation with optional input media.
VideoText to video
Wan 3.0 Prime Uncensored for video generation with optional input media.
VideoText to video
Wan 3.0 Pro Uncensored for video generation with optional input media.
VideoText to video
Wan 3.0 Uncensored for video generation with optional input media.
Text to video · Image to video
Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
Video editing
VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.
Video editing
VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.
Take it into the studio. See the price before you create.
Browse models here without an account. Sign in when you are ready to create.