Unified HappyHorse 1.1 video model. Routes to text-to-video, image-to-video, or reference-to-video with up to 9 reference images based on the inputs you provide.
Added Jun 22, 2026
Approx. Price
$0.420 per video
Model Type
both
Settings
Generation controls available for this model.
Output Format
Default Duration
5
13 duration options
Aspect Ratio
Default
16:9
Options (5)
16:9 (Landscape), 9:16 (Portrait), 1:1 (Square), 4:3 +1 more
Used for text-only generations and optional for reference flows.
Duration
Default
5
Options (13)
3 seconds, 4 seconds, 5 seconds, 6 seconds +9 more
3-15 seconds
Mode
Default
auto
Options (2)
Auto-detect, Reference to Video
Auto-detect uses your uploads. Multiple reference images route to reference-to-video.
Prompt Expansion
Default
true
Options (2)
true, false
Use AI to expand your prompt
Resolution
Default
720p
Options (2)
720p, 1080p
Output resolution
Benchmarks
Benchmarks
Human preference benchmarks sourced from Artificial Analysis.
Text to Video
#6 / 81
ELO
1261.0
Appearances
7,518
95% CI
-8/8
Image to Video
#8 / 75
ELO
1305.0
Appearances
5,153
95% CI
-10/10
Release Date 2026-06 · Matched as HappyHorse-1.1
Artificial Analysis APIExamples
Loading examples…
Related video models
Compare HappyHorse 1.1 with similar models from the same provider or model family.
HappyHorse 1.0
happyhorse-1.0Unified HappyHorse 1.0 video model. Routes to text-to-video, image-to-video, reference-to-video, or video edit based on the inputs you provide.
P-Video Edit
pruna-ai/p-video/editInstruction-based editing for videos up to 15 seconds, with optional reference-image guidance, Draft and Full quality modes, prompt enhancement, and source-audio preservation.
MiniMax H3 Max Turbo
minimax/h3-max-turboMiniMax H3 Max Turbo brings H3 Max prompt understanding and aesthetics to video generation at roughly twice the speed and half the cost while targeting 97% of its quality. Create 5–15 second 480p or 768p clips from text or a starting image, with optional first/last-frame transitions.
MiniMax H3 Spicy Image-to-Video
wavespeed-ai/minimax-h3/image-to-video-spicyOpen-weights MiniMax H3 image-to-video generation with expressive unrestricted motion, native stereo audio, optional last-frame control, 3–15 second clips, and 480p or 768p output.
Gemini Omni Flash 1.1
google/gemini-omni-flash/v1.1Google’s multimodal video model for text-to-video, image animation with optional end frames, multimodal reference generation, and instruction-based video editing. Generates synchronized native audio at resolutions from 360p through 4K.
MiniMax H3 Max
minimax/h3-maxMiniMax H3 Max generates 5–15 second videos at 480p or 768p from a text prompt or a first-frame image, with optional first/last-frame transitions.