AI video models
Every text-to-video and image-to-video model on Cloudflare Workers AI. Open any model to see its exact input schema and an estimated cost per video.
HappyHorse 1.1 · Text
Alibaba
Strong dynamic expressiveness, better visual quality and improved instruction following. Configurable resolution, aspect ratio and duration (3–15s).
HappyHorse 1.0 · Text
Alibaba
HappyHorse 1.0 text-to-video. Generates videos from a text prompt with configurable resolution, aspect ratio and duration (3–15s).
Seedance 2.0
ByteDance
Next-generation video model with synchronized audio. Generates from text, images, video clips and audio. Native audio generation, video editing and extension.
Seedance 2.0 Fast
ByteDance
Faster variant of Seedance 2.0. Trades some quality for speed while sharing the same multimodal architecture.
Seedance 2.0 Mini
ByteDance
Compact, cost-efficient video generation model from the Seedance family. Ideal for high-volume workloads where speed and cost matter.
Seedance 2.5
ByteDance
Audio-video generation model for 30-second videos with reference control and editing capabilities.
FLUX 3 Video
Black Forest Labs
Generates video from a text prompt (t2v), animates one or more reference images (i2v) or continues an existing clip (v2v). Synchronized audio, up to FHD, 5–20s.
Veo 3.1
Google's latest video generation model with improved quality, motion and audio generation.
Veo 3.1 Fast
A faster version of Veo 3.1 optimized for lower latency while maintaining high-quality video and audio output.