Vidu Q1 video generation model. Creates high-quality 5-second videos. Supports both text-to-video and image-to-video generation with customizable visual styles (general or anime), movement amplitude control, and fixed 16:9 output.
Added Jul 10, 2025
Approx. Price
$0.150 per video
Model Type
both
Settings
Generation controls available for this model.
Output Format
N/A
Default Duration
5
Duration
Default
5
Movement Amplitude
Default
auto
Options (4)
Auto, Small, Medium, Large
Amount of movement in the generated video
Size
Default
16:9
Options (1)
16:9 (1920x1080 / Landscape) - Only supported resolution
Video resolution - Vidu only supports 1920x1080 (16:9)
Style
Default
general
Options (2)
General, Anime
Visual style for the video (only available for text-to-video)
Benchmarks
Benchmarks
Human preference benchmarks sourced from Artificial Analysis.
Text to Video
#67 / 81
ELO
1004.0
Appearances
2,914
95% CI
-10/10
Image to Video
#67 / 75
ELO
1023.0
Appearances
3,038
95% CI
-12/12
Release Date 2025-04 · Matched as Vidu Q1
Artificial Analysis APIExamples
Loading examples…
Related video models
Compare Vidu Q1 with similar models from the same provider or model family.
Vidu Q3 Pro
vidu-q3-proVidu Q3 Pro text-to-video, image-to-video, and start/end-frame video generation with high visual fidelity, 540p/720p/1080p output, 1-16s duration, and optional audio plus background music.
Vidu Q3
vidu-q3Vidu Q3 text-to-video and image-to-video with high visual fidelity, multiple styles, 540p/720p/1080p output, 1-16s duration, and optional audio plus background music.
P-Video Edit
pruna-ai/p-video/editInstruction-based editing for videos up to 15 seconds, with optional reference-image guidance, Draft and Full quality modes, prompt enhancement, and source-audio preservation.
MiniMax H3 Max Turbo
minimax/h3-max-turboMiniMax H3 Max Turbo brings H3 Max prompt understanding and aesthetics to video generation at roughly twice the speed and half the cost while targeting 97% of its quality. Create 5–15 second 480p or 768p clips from text or a starting image, with optional first/last-frame transitions.
MiniMax H3 Spicy Image-to-Video
wavespeed-ai/minimax-h3/image-to-video-spicyOpen-weights MiniMax H3 image-to-video generation with expressive unrestricted motion, native stereo audio, optional last-frame control, 3–15 second clips, and 480p or 768p output.
Gemini Omni Flash 1.1
google/gemini-omni-flash/v1.1Google’s multimodal video model for text-to-video, image animation with optional end frames, multimodal reference generation, and instruction-based video editing. Generates synchronized native audio at resolutions from 360p through 4K.