Happy Horse 1.1
by Alibaba
Alibaba's video model.
Synchronized audio and native lip‑sync, generated in a single pass.

Key features
Technical specifications
1080p
Render at 1080p for delivery, or 720p to draft faster.
3–15s
Each clip runs 3 to 15 seconds, with a 5-second default.
7
Native lip-sync in seven languages, matched to each one's phonetics.
Up to 9
Bring up to nine subjects, each called by index in the prompt.
Use cases
Dialogue-driven scenes
Characters speak in any of 7 languages with synced lip movement, ambient sound, and timing, generated together in one pass.
Multi-character storytelling
Hold up to nine subjects from reference images and carry them across scenes, calling each by index for consistent ensemble work.
Ad and campaign spots
Reference-driven control keeps product, talent, and brand visuals consistent across shots, with audio and motion in sync.
Music videos and performance
Video and audio generated together means motion lands on beat from the first pass, with no manual sync work afterward.
Ultrawide and vertical
Deliver the same scene as a 21:9 cinematic cut and a 9:16 vertical from nine aspect ratios, no separate workflow per format.
Multilingual localization
Same scene, same characters, dialogue swapped across languages with native lip-sync, suited for global campaigns.
Prompt examples
Overview
Happy Horse 1.1 is Alibaba's video model, served on fal and available on Morphic. It generates video and audio jointly in a single pass, producing 1080p clips of 3 to 15 seconds with native lip-sync across seven languages. Its reference-driven control across text, image, and reference inputs suits dialogue scenes, performance clips, and character-consistent narrative work.
Reference-to-video and delivery
Happy Horse 1.1 carries up to nine reference subjects into a scene, each called by index from character1 to character9, which opens up ensemble and multi-character work. It delivers in nine aspect ratios, from 16:9 and 9:16 to ultrawide 21:9, plus 9:21, 5:4, and 4:5.
Happy Horse 1.1 on Morphic
On Morphic, Happy Horse 1.1 sits in the video model picker alongside Seedance 2.0, Veo, Kling, and the rest of the video catalog. Switch the prompt bar to Video mode, pick Happy Horse 1.1, attach a still or up to nine reference images, and generate a clip with synchronized audio in one pass.
Simple pricing
Get started for free today, with the option to upgrade or cancel anytime.
Basic
1100 monthly credits
1 user only
All models
Workflows
Standard
3625 monthly credits
1 user only
All models
Workflows
Pro
6350 shared monthly credits
1 user
All models
Workflows
Pro Max
24650 shared monthly credits
1 user
All models
Workflows
Enterprise
For higher limits
Custom
pricing and billing terms

Free
For playing around
$0
forever free
FAQs
ChatGPT Images 2.5
OpenAI
OpenAI's image model, released September 2026. Sharper detail, precise edits, faster.
Lyria 3.5
Google DeepMind
Google's best-sounding music model. Full songs with structure, vocals, and lyrics.
MiniMax H3 Max Turbo
fal.ai
The fast tier of fal's post-trained H3. Twice the speed, nearly the same look.
Inworld TTS 2
Inworld
Ninety-five voices, 100+ languages. Speech that starts in under a second.