Skip to main content

Video models

A reviewed catalog of video models for text, photo, and reference generation; Neiron availability is marked separately.

Seedance 2.5

ByteDance

A fast video model for clips, ad scenes, and visual idea tests.

Video modelsAvailable in NeironText-to-video generationReference-image workflows

Veo 3.1

Google

Google video model for expressive scenes, camera motion, and clips with audio context.

Video modelsAvailable in NeironShort prompt-to-video clipsScenes with audio and speech

Gemini Omni

Google

A model for multimodal video scenes where references and prompt following matter.

Video modelsAvailable in NeironPrompt-to-videoReference support

Grok Imagine

xAI

A creative video model for fast ideas, meme-like scenes, and unusual visual moves.

Video modelsAvailable in NeironFast video generationCreative scenes

Wan 2.6

Wan

A practical model for video-first tasks that need different frame formats.

Video modelsAvailable in NeironFlexible frame formatsImage-to-video

Kling Motion

Kling AI

A model for motion templates, dance clips, and animating photos.

Video modelsAvailable in NeironPhoto animationMotion templates

Grok Imagine Text to Video

xAI

Grok Imagine Text to Video by xAI can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Grok Imagine Image to Video

xAI

Grok Imagine Image to Video by xAI can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Grok Imagine Video Upscale

xAI

Grok Imagine Video Upscale by xAI can create or transform video. Input: Video. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video Upscale: Create or transform video.Input: Video; output: Video.

Grok Imagine Video Extend

xAI

Grok Imagine Video Extend by xAI can extend an existing video or music segment. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video Extend: Extend an existing video or music segment.Input: Text, Video; output: Video.

Grok Imagine Video 1.5 Preview

xAI

Grok Imagine Video 1.5 Preview by xAI can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video 1.5 Preview: Create or transform video.Input: Text, Images; output: Video.

Kling 2.6 Text to Video

Kuaishou

Kling 2.6 Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 2.6 Image to Video

Kuaishou

Kling 2.6 Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.5 Turbo Pro Image to Video

Kuaishou

Kling 2.5 Turbo Pro Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.5 Turbo Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.5 Turbo Pro Text to Video

Kuaishou

Kling 2.5 Turbo Pro Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.5 Turbo Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling AI Avatar Standard

Kuaishou

Kling AI Avatar Standard by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageKling AI Avatar Standard: Create or analyze a digital human and its presence in the frame.Input: Text, Images, Audio; output: Video.

Kling AI Avatar Pro

Kuaishou

Kling AI Avatar Pro by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageKling AI Avatar Pro: Create or analyze a digital human and its presence in the frame.Input: Text, Images, Audio; output: Video.

Kling 2.1 Master Image to Video

Kuaishou

Kling 2.1 Master Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Master Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.1 Master Text to Video

Kuaishou

Kling 2.1 Master Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Master Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 2.1 Pro

Kuaishou

Kling 2.1 Pro by Kuaishou can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Pro: Create or transform video.Input: Text, Images; output: Video.

Kling 2.1 Standard

Kuaishou

Kling 2.1 Standard by Kuaishou can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Standard: Create or transform video.Input: Text, Images; output: Video.

Kling 2.6 Motion Control

Kuaishou

Kling 2.6 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Motion Control: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Kling 3.0 Motion Control

Kuaishou

Kling 3.0 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageKling 3.0 Motion Control: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Kling 3.0

Kuaishou

Kling 3.0 by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 3.0: Create a video scene from a text brief.Input: Text; output: Video.

Kling 3 Turbo Text to Video

Kuaishou

Kling 3 Turbo Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 3 Turbo Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 3 Turbo Image to Video

Kuaishou

Kling 3 Turbo Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 3 Turbo Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Seedance 2.0 Fast

ByteDance

Seedance 2.0 Fast by ByteDance can create a video scene from a text brief. Input: Text, Images, Video, Audio. Output: Video.

Video modelsExternal model, reference pageSeedance 2.0 Fast: Create a video scene from a text brief.Input: Text, Images, Video, Audio; output: Video.

Seedance 2.0 Mini

ByteDance

Seedance 2.0 Mini by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageSeedance 2.0 Mini: Create a video scene from a text brief.Input: Text, Images; output: Video.

Seedance 1.5 Pro

ByteDance

Seedance 1.5 Pro by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageSeedance 1.5 Pro: Create a video scene from a text brief.Input: Text, Images; output: Video.

ByteDance V1 Pro Fast Image to Video

ByteDance

ByteDance V1 Pro Fast Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Fast Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Pro Image to Video

ByteDance

ByteDance V1 Pro Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Pro Text to Video

ByteDance

ByteDance V1 Pro Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

ByteDance V1 Lite Image to Video

ByteDance

ByteDance V1 Lite Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Lite Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Lite Text to Video

ByteDance

ByteDance V1 Lite Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Lite Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo 2.3 Pro Image to Video

MiniMax

Hailuo 2.3 Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo 2.3 Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo 2.3 Standard Image to Video

MiniMax

Hailuo 2.3 Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo 2.3 Standard Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo Pro Text to Video

MiniMax

Hailuo Pro Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHailuo Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo Pro Image to Video

MiniMax

Hailuo Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo Standard Text to Video

MiniMax

Hailuo Standard Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHailuo Standard Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo Standard Image to Video

MiniMax

Hailuo Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo Standard Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.2 A14B Image to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Image to Video Turbo by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Image to Video Turbo: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.2 A14B Speech to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Speech to Video Turbo by Alibaba Cloud can create video, motion, or character speech from an audio recording. Input: Text, Audio, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Speech to Video Turbo: Create video, motion, or character speech from an audio recording.Input: Text, Audio, Images; output: Video.

Wan 2.2 A14B Text to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Text to Video Turbo by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Text to Video Turbo: Create a video scene from a text brief.Input: Text; output: Video.

Wan Animate Move

Alibaba Cloud

Wan Animate Move by Alibaba Cloud can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageWan Animate Move: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Wan Animate Replace

Alibaba Cloud

Wan Animate Replace by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageWan Animate Replace: Transform an existing video without reshooting the source scene.Input: Images, Video; output: Video.

Wan 2.6 Image to Video

Alibaba Cloud

Wan 2.6 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.6 Text to Video

Alibaba Cloud

Wan 2.6 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.6 Video to Video

Alibaba Cloud

Wan 2.6 Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Video to Video: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.6 Flash Image to Video

Alibaba Cloud

Wan 2.6 Flash Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Flash Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.6 Flash Video to Video

Alibaba Cloud

Wan 2.6 Flash Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Flash Video to Video: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.5 Image to Video

Alibaba Cloud

Wan 2.5 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.5 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.5 Text to Video

Alibaba Cloud

Wan 2.5 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.5 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.7 Text to Video

Alibaba Cloud

Wan 2.7 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.7 Image to Video

Alibaba Cloud

Wan 2.7 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.7 Video Edit

Alibaba Cloud

Wan 2.7 Video Edit by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Video Edit: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.7 Reference to Video

Alibaba Cloud

Wan 2.7 Reference to Video by Alibaba Cloud can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

Topaz Video Upscale

Topaz Labs

Topaz Video Upscale by Topaz Labs can create or transform video. Input: Video. Output: Video.

Video modelsExternal model, reference pageTopaz Video Upscale: Create or transform video.Input: Video; output: Video.

Infinitalk From Audio

MeiGen AI

Infinitalk From Audio by MeiGen AI can create or analyze a digital human and its presence in the frame. Input: Audio, Images. Output: Video.

Video modelsExternal model, reference pageInfinitalk From Audio: Create or analyze a digital human and its presence in the frame.Input: Audio, Images; output: Video.

Runway Aleph

Runway

Runway Aleph by Runway can create or transform video. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageRunway Aleph: Create or transform video.Input: Text, Video; output: Video.

Runway AI Video

Runway

Runway AI Video by Runway can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageRunway AI Video: Create a video scene from a text brief.Input: Text, Images; output: Video.

Runway Video Extend

Runway

Runway Video Extend by Runway can extend an existing video or music segment. Input: Video. Output: Video.

Video modelsExternal model, reference pageRunway Video Extend: Extend an existing video or music segment.Input: Video; output: Video.

PixVerse V6 Text to Video

PixVerse

PixVerse V6 Text to Video by PixVerse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

PixVerse V6 Image to Video

PixVerse

PixVerse V6 Image to Video by PixVerse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

PixVerse V6 First and Last Frame Transition

PixVerse

PixVerse V6 First and Last Frame Transition by PixVerse can construct a transition between supplied first and last frames. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 First and Last Frame Transition: Construct a transition between supplied first and last frames.Input: Text, Images; output: Video.

PixVerse V6 Video Extension

PixVerse

PixVerse V6 Video Extension by PixVerse can extend an existing video or music segment. Input: Video. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Video Extension: Extend an existing video or music segment.Input: Video; output: Video.

PixVerse V6 Fusion Reference to Video

PixVerse

PixVerse V6 Fusion Reference to Video by PixVerse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Fusion Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

MiniMax H3 Text to Video

MiniMax

MiniMax H3 Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

MiniMax H3 Image to Video

MiniMax

MiniMax H3 Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

MiniMax H3 Reference to Video

MiniMax

MiniMax H3 Reference to Video by MiniMax can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

HappyHorse Text to Video

HappyHorse

HappyHorse Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHappyHorse Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

HappyHorse Image to Video

HappyHorse

HappyHorse Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHappyHorse Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

HappyHorse Reference to Video

HappyHorse

HappyHorse Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

HappyHorse Video Edit

HappyHorse

HappyHorse Video Edit by HappyHorse can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse Video Edit: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

HappyHorse 1.1 Image to Video

HappyHorse

HappyHorse 1.1 Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

HappyHorse 1.1 Text to Video

HappyHorse

HappyHorse 1.1 Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

HappyHorse 1.1 Reference to Video

HappyHorse

HappyHorse 1.1 Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

Gemini Omni Video

Google

Gemini Omni Video by Google can create or transform video. Input: Text, Images, Video, Audio. Output: Video.

Video modelsExternal model, reference pageGemini Omni Video: Create or transform video.Input: Text, Images, Video, Audio; output: Video.

Gemini Omni Character

Google

Gemini Omni Character by Google can create or transform video. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageGemini Omni Character: Create or transform video.Input: Text, Images, Audio; output: Video.

OmniHuman 1.5

ByteDance

OmniHuman 1.5 by ByteDance can create or analyze a digital human and its presence in the frame. Input: Images, Audio, Video. Output: Video.

Video modelsExternal model, reference pageOmniHuman 1.5: Create or analyze a digital human and its presence in the frame.Input: Images, Audio, Video; output: Video.

Volcengine Video Lip Sync

ByteDance

Volcengine Video Lip Sync by ByteDance can synchronize lip movement in video with an audio track. Input: Video, Audio. Output: Video.

Video modelsExternal model, reference pageVolcengine Video Lip Sync: Synchronize lip movement in video with an audio track.Input: Video, Audio; output: Video.

Suno Music Video

Suno

Suno Music Video by Suno can create, transform, or continue musical material. Input: Audio, Images. Output: Video.

Video modelsExternal model, reference pageSuno Music Video: Create, transform, or continue musical material.Input: Audio, Images; output: Video.

How to choose a video model

This page groups models of one type, but they solve different jobs. Start from the use case, not the model name: it is the shortest path to a usable result without unnecessary reruns.

  • animate a photo or create a short clip from a text brief
  • pick a model by format, duration, and motion style
  • test several variants before publishing to social channels or ads

Practical choice

  • Seedance 2.5 is the fast mode for clips, ad scenes, and visual idea tests.
  • Veo 3.1 supports scenes with audio and speech, while Gemini Omni focuses on references and prompt following.
  • Kling Motion fits vertical video and controlled-motion scenarios.