Video models
A reviewed catalog of video models for text, photo, and reference generation; Neiron availability is marked separately.
Seedance 2.5
ByteDance
A fast video model for clips, ad scenes, and visual idea tests.
Veo 3.1
Google video model for expressive scenes, camera motion, and clips with audio context.
Gemini Omni
A model for multimodal video scenes where references and prompt following matter.
Grok Imagine
xAI
A creative video model for fast ideas, meme-like scenes, and unusual visual moves.
Wan 2.6
Wan
A practical model for video-first tasks that need different frame formats.
Kling Motion
Kling AI
A model for motion templates, dance clips, and animating photos.
Grok Imagine Text to Video
xAI
Grok Imagine Text to Video by xAI can create a video scene from a text brief. Input: Text. Output: Video.
Grok Imagine Image to Video
xAI
Grok Imagine Image to Video by xAI can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Grok Imagine Video Upscale
xAI
Grok Imagine Video Upscale by xAI can create or transform video. Input: Video. Output: Video.
Grok Imagine Video Extend
xAI
Grok Imagine Video Extend by xAI can extend an existing video or music segment. Input: Text, Video. Output: Video.
Grok Imagine Video 1.5 Preview
xAI
Grok Imagine Video 1.5 Preview by xAI can create or transform video. Input: Text, Images. Output: Video.
Kling 2.6 Text to Video
Kuaishou
Kling 2.6 Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.
Kling 2.6 Image to Video
Kuaishou
Kling 2.6 Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Kling 2.5 Turbo Pro Image to Video
Kuaishou
Kling 2.5 Turbo Pro Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Kling 2.5 Turbo Pro Text to Video
Kuaishou
Kling 2.5 Turbo Pro Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.
Kling AI Avatar Standard
Kuaishou
Kling AI Avatar Standard by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.
Kling AI Avatar Pro
Kuaishou
Kling AI Avatar Pro by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.
Kling 2.1 Master Image to Video
Kuaishou
Kling 2.1 Master Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Kling 2.1 Master Text to Video
Kuaishou
Kling 2.1 Master Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.
Kling 2.1 Pro
Kuaishou
Kling 2.1 Pro by Kuaishou can create or transform video. Input: Text, Images. Output: Video.
Kling 2.1 Standard
Kuaishou
Kling 2.1 Standard by Kuaishou can create or transform video. Input: Text, Images. Output: Video.
Kling 2.6 Motion Control
Kuaishou
Kling 2.6 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.
Kling 3.0 Motion Control
Kuaishou
Kling 3.0 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.
Kling 3.0
Kuaishou
Kling 3.0 by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.
Kling 3 Turbo Text to Video
Kuaishou
Kling 3 Turbo Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.
Kling 3 Turbo Image to Video
Kuaishou
Kling 3 Turbo Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Seedance 2.0 Fast
ByteDance
Seedance 2.0 Fast by ByteDance can create a video scene from a text brief. Input: Text, Images, Video, Audio. Output: Video.
Seedance 2.0 Mini
ByteDance
Seedance 2.0 Mini by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.
Seedance 1.5 Pro
ByteDance
Seedance 1.5 Pro by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.
ByteDance V1 Pro Fast Image to Video
ByteDance
ByteDance V1 Pro Fast Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
ByteDance V1 Pro Image to Video
ByteDance
ByteDance V1 Pro Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
ByteDance V1 Pro Text to Video
ByteDance
ByteDance V1 Pro Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.
ByteDance V1 Lite Image to Video
ByteDance
ByteDance V1 Lite Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
ByteDance V1 Lite Text to Video
ByteDance
ByteDance V1 Lite Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.
Hailuo 2.3 Pro Image to Video
MiniMax
Hailuo 2.3 Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Hailuo 2.3 Standard Image to Video
MiniMax
Hailuo 2.3 Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Hailuo Pro Text to Video
MiniMax
Hailuo Pro Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.
Hailuo Pro Image to Video
MiniMax
Hailuo Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Hailuo Standard Text to Video
MiniMax
Hailuo Standard Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.
Hailuo Standard Image to Video
MiniMax
Hailuo Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.2 A14B Image to Video Turbo
Alibaba Cloud
Wan 2.2 A14B Image to Video Turbo by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.2 A14B Speech to Video Turbo
Alibaba Cloud
Wan 2.2 A14B Speech to Video Turbo by Alibaba Cloud can create video, motion, or character speech from an audio recording. Input: Text, Audio, Images. Output: Video.
Wan 2.2 A14B Text to Video Turbo
Alibaba Cloud
Wan 2.2 A14B Text to Video Turbo by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.
Wan Animate Move
Alibaba Cloud
Wan Animate Move by Alibaba Cloud can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.
Wan Animate Replace
Alibaba Cloud
Wan Animate Replace by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Images, Video. Output: Video.
Wan 2.6 Image to Video
Alibaba Cloud
Wan 2.6 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.6 Text to Video
Alibaba Cloud
Wan 2.6 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.
Wan 2.6 Video to Video
Alibaba Cloud
Wan 2.6 Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.
Wan 2.6 Flash Image to Video
Alibaba Cloud
Wan 2.6 Flash Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.6 Flash Video to Video
Alibaba Cloud
Wan 2.6 Flash Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.
Wan 2.5 Image to Video
Alibaba Cloud
Wan 2.5 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.5 Text to Video
Alibaba Cloud
Wan 2.5 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.
Wan 2.7 Text to Video
Alibaba Cloud
Wan 2.7 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.
Wan 2.7 Image to Video
Alibaba Cloud
Wan 2.7 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
Wan 2.7 Video Edit
Alibaba Cloud
Wan 2.7 Video Edit by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.
Wan 2.7 Reference to Video
Alibaba Cloud
Wan 2.7 Reference to Video by Alibaba Cloud can build video from one or more visual references. Input: Text, Images, Video. Output: Video.
Topaz Video Upscale
Topaz Labs
Topaz Video Upscale by Topaz Labs can create or transform video. Input: Video. Output: Video.
Infinitalk From Audio
MeiGen AI
Infinitalk From Audio by MeiGen AI can create or analyze a digital human and its presence in the frame. Input: Audio, Images. Output: Video.
Runway Aleph
Runway
Runway Aleph by Runway can create or transform video. Input: Text, Video. Output: Video.
Runway AI Video
Runway
Runway AI Video by Runway can create a video scene from a text brief. Input: Text, Images. Output: Video.
Runway Video Extend
Runway
Runway Video Extend by Runway can extend an existing video or music segment. Input: Video. Output: Video.
PixVerse V6 Text to Video
PixVerse
PixVerse V6 Text to Video by PixVerse can create a video scene from a text brief. Input: Text. Output: Video.
PixVerse V6 Image to Video
PixVerse
PixVerse V6 Image to Video by PixVerse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
PixVerse V6 First and Last Frame Transition
PixVerse
PixVerse V6 First and Last Frame Transition by PixVerse can construct a transition between supplied first and last frames. Input: Text, Images. Output: Video.
PixVerse V6 Video Extension
PixVerse
PixVerse V6 Video Extension by PixVerse can extend an existing video or music segment. Input: Video. Output: Video.
PixVerse V6 Fusion Reference to Video
PixVerse
PixVerse V6 Fusion Reference to Video by PixVerse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.
MiniMax H3 Text to Video
MiniMax
MiniMax H3 Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.
MiniMax H3 Image to Video
MiniMax
MiniMax H3 Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
MiniMax H3 Reference to Video
MiniMax
MiniMax H3 Reference to Video by MiniMax can build video from one or more visual references. Input: Text, Images, Video. Output: Video.
HappyHorse Text to Video
HappyHorse
HappyHorse Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.
HappyHorse Image to Video
HappyHorse
HappyHorse Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
HappyHorse Reference to Video
HappyHorse
HappyHorse Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.
HappyHorse Video Edit
HappyHorse
HappyHorse Video Edit by HappyHorse can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.
HappyHorse 1.1 Image to Video
HappyHorse
HappyHorse 1.1 Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.
HappyHorse 1.1 Text to Video
HappyHorse
HappyHorse 1.1 Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.
HappyHorse 1.1 Reference to Video
HappyHorse
HappyHorse 1.1 Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.
Gemini Omni Video
Gemini Omni Video by Google can create or transform video. Input: Text, Images, Video, Audio. Output: Video.
Gemini Omni Character
Gemini Omni Character by Google can create or transform video. Input: Text, Images, Audio. Output: Video.
OmniHuman 1.5
ByteDance
OmniHuman 1.5 by ByteDance can create or analyze a digital human and its presence in the frame. Input: Images, Audio, Video. Output: Video.
Volcengine Video Lip Sync
ByteDance
Volcengine Video Lip Sync by ByteDance can synchronize lip movement in video with an audio track. Input: Video, Audio. Output: Video.
Suno Music Video
Suno
Suno Music Video by Suno can create, transform, or continue musical material. Input: Audio, Images. Output: Video.
How to choose a video model
This page groups models of one type, but they solve different jobs. Start from the use case, not the model name: it is the shortest path to a usable result without unnecessary reruns.
- animate a photo or create a short clip from a text brief
- pick a model by format, duration, and motion style
- test several variants before publishing to social channels or ads
Practical choice
- Seedance 2.5 is the fast mode for clips, ad scenes, and visual idea tests.
- Veo 3.1 supports scenes with audio and speech, while Gemini Omni focuses on references and prompt following.
- Kling Motion fits vertical video and controlled-motion scenarios.