Vidu Q1 video generation model. Creates high-quality 5-second videos. Supports both text-to-video and image-to-video generation with customizable visual styles (general or anime), movement amplitude control, and fixed 16:9 output.
Added Jul 10, 2025
Approx. Price
$0.150 per video
Model Type
both
Settings
Generation controls available for this model.
Output Format
N/A
Default Duration
5
Duration
Default
5
Movement Amplitude
Default
auto
Options (4)
Auto, Small, Medium, Large
Amount of movement in the generated video
Size
Default
16:9
Options (1)
16:9 (1920x1080 / Landscape) - Only supported resolution
Video resolution - Vidu only supports 1920x1080 (16:9)
Style
Default
general
Options (2)
General, Anime
Visual style for the video (only available for text-to-video)
Benchmarks
Benchmarks
Human preference benchmarks sourced from Artificial Analysis.
Text to Video
#64 / 78
ELO
1004.0
Appearances
2,915
95% CI
-10/10
Image to Video
#64 / 72
ELO
1023.0
Appearances
3,037
95% CI
-12/12
Release Date 2025-04 · Matched as Vidu Q1
Artificial Analysis APIExamples
Loading examples…
Related video models
Compare Vidu Q1 with similar models from the same provider or model family.
Vidu Q3 Pro
vidu-q3-proVidu Q3 Pro text-to-video, image-to-video, and start/end-frame video generation with high visual fidelity, 540p/720p/1080p output, 1-16s duration, and optional audio plus background music.
Vidu Q3
vidu-q3Vidu Q3 text-to-video and image-to-video with high visual fidelity, multiple styles, 540p/720p/1080p output, 1-16s duration, and optional audio plus background music.
LTX-2.5 Fast
lightricks/ltx-2.5/fastSpeed-optimized audiovisual generation from text, an image, or a 2-20 second audio clip. Creates synchronized video and audio in one pass, with output up to 4K and optional start/end-frame control.
LTX-2.5 Pro
lightricks/ltx-2.5/proHigh-fidelity audiovisual generation from text, an image, or a 2-20 second audio clip. Creates polished synchronized video and audio in one pass, with 720p/1080p output and optional start/end-frame control.
Wan 3.0 Image-to-Video
alibaba/wan-3.0/image-to-videoAnimate a first-frame image into a cinematic video with optional last-frame guidance, synchronized audio, deep-thinking controls, and 2–30 second output.
Wan 3.0 Reference-to-Video
alibaba/wan-3.0/reference-to-videoReference-guided video generation using images, videos, and audio for subject consistency, motion, timing, and scene continuity, with 2–30 second output.