Skip to main content

AI models: catalog and comparison

Compare text models, image generators, and video models by task, input, and mode. The catalog clearly separates models available in Neiron now from external market reference cards.

20 models are currently available in Neiron. Check each model card for the current selection, restrictions, and allowance usage.

View AI access plans
Output
Input
Purpose
Developer
Geography

Models and modes found: 207

Video models

A reviewed catalog of video models for text, photo, and reference generation; Neiron availability is marked separately.

All models and modes

Seedance 2.5

ByteDance

Describe one scene: who is in frame, what happens, and how the camera moves. Seedance can create a clip from text or photo references. Format, duration, and reference limits depend on the variant selected in the generator; check them before running.

Video modelsAvailable in NeironBuild a video scene from action and camera instructions.Guide a scene or subject with photo references.

Veo 3.1

Google

With Veo, describe the scene together with sound, such as rain, street noise, or a short spoken line. Specify visual action and the audio setting separately. Review the full clip afterward: an audio track does not guarantee the requested wording or synchronization.

Video modelsAvailable in NeironCreate a short scene from a text brief.Describe the audio setting in the prompt.

Gemini Omni

Google

Give Omni both a description and source material: photos or a video. Explain what each should contribute, such as the subject, setting, or movement. A request containing source video allows a different number of photo references from a request without it.

Video modelsAvailable in NeironCreate a clip from a scene description.Use photo references together with a written brief.

Grok Imagine

xAI

Grok Imagine here generates a video rather than answering a chat question. Describe a scene or use a photo as the starting point. For an initial prompt, specify one visible movement: an object turning, fabric moving, or the camera approaching.

Video modelsAvailable in NeironCreate a short clip from a written brief.Add motion to a scene guided by a photo.

Wan 2.6

Wan

Start a Wan prompt with the source scene and a small change over time: camera movement, lighting, or an object’s action. The current generator offers square, portrait, and landscape frames. Select the aspect ratio in the panel rather than only writing it in the prompt.

Video modelsAvailable in NeironTurn a photo into a video scene with requested motion.Generate a clip from a written description.

Kling Motion

Kling AI

Here a template or source motion video provides movement, while a photo provides appearance. Choose an image where the figure is visible and not obscured by large objects. This is motion transfer, not unrestricted scene generation from text alone.

Video modelsAvailable in NeironAnimate a photographed character using selected motion.Use a preset template or your own motion clip.

Grok Imagine Text to Video

xAI

Grok Imagine Text to Video by xAI can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Grok Imagine Image to Video

xAI

Grok Imagine Image to Video by xAI can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Grok Imagine Video Upscale

xAI

Grok Imagine Video Upscale by xAI can create or transform video. Input: Video. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video Upscale: Create or transform video.Input: Video; output: Video.

Grok Imagine Video Extend

xAI

Grok Imagine Video Extend by xAI can extend an existing video or music segment. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video Extend: Extend an existing video or music segment.Input: Text, Video; output: Video.

Grok Imagine Video 1.5 Preview

xAI

Grok Imagine Video 1.5 Preview by xAI can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageGrok Imagine Video 1.5 Preview: Create or transform video.Input: Text, Images; output: Video.

Kling 2.6 Text to Video

Kuaishou

Kling 2.6 Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 2.6 Image to Video

Kuaishou

Kling 2.6 Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.5 Turbo Pro Image to Video

Kuaishou

Kling 2.5 Turbo Pro Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.5 Turbo Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.5 Turbo Pro Text to Video

Kuaishou

Kling 2.5 Turbo Pro Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.5 Turbo Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling AI Avatar Standard

Kuaishou

Kling AI Avatar Standard by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageKling AI Avatar Standard: Create or analyze a digital human and its presence in the frame.Input: Text, Images, Audio; output: Video.

Kling AI Avatar Pro

Kuaishou

Kling AI Avatar Pro by Kuaishou can create or analyze a digital human and its presence in the frame. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageKling AI Avatar Pro: Create or analyze a digital human and its presence in the frame.Input: Text, Images, Audio; output: Video.

Kling 2.1 Master Image to Video

Kuaishou

Kling 2.1 Master Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Master Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Kling 2.1 Master Text to Video

Kuaishou

Kling 2.1 Master Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Master Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 2.1 Pro

Kuaishou

Kling 2.1 Pro by Kuaishou can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Pro: Create or transform video.Input: Text, Images; output: Video.

Kling 2.1 Standard

Kuaishou

Kling 2.1 Standard by Kuaishou can create or transform video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 2.1 Standard: Create or transform video.Input: Text, Images; output: Video.

Kling 2.6 Motion Control

Kuaishou

Kling 2.6 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageKling 2.6 Motion Control: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Kling 3.0 Motion Control

Kuaishou

Kling 3.0 Motion Control by Kuaishou can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageKling 3.0 Motion Control: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Kling 3.0

Kuaishou

Kling 3.0 by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 3.0: Create a video scene from a text brief.Input: Text; output: Video.

Kling 3 Turbo Text to Video

Kuaishou

Kling 3 Turbo Text to Video by Kuaishou can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageKling 3 Turbo Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Kling 3 Turbo Image to Video

Kuaishou

Kling 3 Turbo Image to Video by Kuaishou can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageKling 3 Turbo Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Seedance 2.0 Fast

ByteDance

Seedance 2.0 Fast by ByteDance can create a video scene from a text brief. Input: Text, Images, Video, Audio. Output: Video.

Video modelsExternal model, reference pageSeedance 2.0 Fast: Create a video scene from a text brief.Input: Text, Images, Video, Audio; output: Video.

Seedance 2.0 Mini

ByteDance

Seedance 2.0 Mini by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageSeedance 2.0 Mini: Create a video scene from a text brief.Input: Text, Images; output: Video.

Seedance 1.5 Pro

ByteDance

Seedance 1.5 Pro by ByteDance can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageSeedance 1.5 Pro: Create a video scene from a text brief.Input: Text, Images; output: Video.

ByteDance V1 Pro Fast Image to Video

ByteDance

ByteDance V1 Pro Fast Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Fast Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Pro Image to Video

ByteDance

ByteDance V1 Pro Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Pro Text to Video

ByteDance

ByteDance V1 Pro Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

ByteDance V1 Lite Image to Video

ByteDance

ByteDance V1 Lite Image to Video by ByteDance can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Lite Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

ByteDance V1 Lite Text to Video

ByteDance

ByteDance V1 Lite Text to Video by ByteDance can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageByteDance V1 Lite Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo 2.3 Pro Image to Video

MiniMax

Hailuo 2.3 Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo 2.3 Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo 2.3 Standard Image to Video

MiniMax

Hailuo 2.3 Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo 2.3 Standard Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo Pro Text to Video

MiniMax

Hailuo Pro Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHailuo Pro Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo Pro Image to Video

MiniMax

Hailuo Pro Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo Pro Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Hailuo Standard Text to Video

MiniMax

Hailuo Standard Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHailuo Standard Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Hailuo Standard Image to Video

MiniMax

Hailuo Standard Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHailuo Standard Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.2 A14B Image to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Image to Video Turbo by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Image to Video Turbo: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.2 A14B Speech to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Speech to Video Turbo by Alibaba Cloud can create video, motion, or character speech from an audio recording. Input: Text, Audio, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Speech to Video Turbo: Create video, motion, or character speech from an audio recording.Input: Text, Audio, Images; output: Video.

Wan 2.2 A14B Text to Video Turbo

Alibaba Cloud

Wan 2.2 A14B Text to Video Turbo by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.2 A14B Text to Video Turbo: Create a video scene from a text brief.Input: Text; output: Video.

Wan Animate Move

Alibaba Cloud

Wan Animate Move by Alibaba Cloud can control character or camera motion from a supplied reference. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageWan Animate Move: Control character or camera motion from a supplied reference.Input: Images, Video; output: Video.

Wan Animate Replace

Alibaba Cloud

Wan Animate Replace by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Images, Video. Output: Video.

Video modelsExternal model, reference pageWan Animate Replace: Transform an existing video without reshooting the source scene.Input: Images, Video; output: Video.

Wan 2.6 Image to Video

Alibaba Cloud

Wan 2.6 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.6 Text to Video

Alibaba Cloud

Wan 2.6 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.6 Video to Video

Alibaba Cloud

Wan 2.6 Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Video to Video: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.6 Flash Image to Video

Alibaba Cloud

Wan 2.6 Flash Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Flash Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.6 Flash Video to Video

Alibaba Cloud

Wan 2.6 Flash Video to Video by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.6 Flash Video to Video: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.5 Image to Video

Alibaba Cloud

Wan 2.5 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.5 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.5 Text to Video

Alibaba Cloud

Wan 2.5 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.5 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.7 Text to Video

Alibaba Cloud

Wan 2.7 Text to Video by Alibaba Cloud can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

Wan 2.7 Image to Video

Alibaba Cloud

Wan 2.7 Image to Video by Alibaba Cloud can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

Wan 2.7 Video Edit

Alibaba Cloud

Wan 2.7 Video Edit by Alibaba Cloud can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Video Edit: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

Wan 2.7 Reference to Video

Alibaba Cloud

Wan 2.7 Reference to Video by Alibaba Cloud can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageWan 2.7 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

Topaz Video Upscale

Topaz Labs

Topaz Video Upscale by Topaz Labs can create or transform video. Input: Video. Output: Video.

Video modelsExternal model, reference pageTopaz Video Upscale: Create or transform video.Input: Video; output: Video.

Infinitalk From Audio

MeiGen AI

Infinitalk From Audio by MeiGen AI can create or analyze a digital human and its presence in the frame. Input: Audio, Images. Output: Video.

Video modelsExternal model, reference pageInfinitalk From Audio: Create or analyze a digital human and its presence in the frame.Input: Audio, Images; output: Video.

Runway Aleph

Runway

Runway Aleph by Runway can create or transform video. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageRunway Aleph: Create or transform video.Input: Text, Video; output: Video.

Runway AI Video

Runway

Runway AI Video by Runway can create a video scene from a text brief. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageRunway AI Video: Create a video scene from a text brief.Input: Text, Images; output: Video.

Runway Video Extend

Runway

Runway Video Extend by Runway can extend an existing video or music segment. Input: Video. Output: Video.

Video modelsExternal model, reference pageRunway Video Extend: Extend an existing video or music segment.Input: Video; output: Video.

PixVerse V6 Text to Video

PixVerse

PixVerse V6 Text to Video by PixVerse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

PixVerse V6 Image to Video

PixVerse

PixVerse V6 Image to Video by PixVerse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

PixVerse V6 First and Last Frame Transition

PixVerse

PixVerse V6 First and Last Frame Transition by PixVerse can construct a transition between supplied first and last frames. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 First and Last Frame Transition: Construct a transition between supplied first and last frames.Input: Text, Images; output: Video.

PixVerse V6 Video Extension

PixVerse

PixVerse V6 Video Extension by PixVerse can extend an existing video or music segment. Input: Video. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Video Extension: Extend an existing video or music segment.Input: Video; output: Video.

PixVerse V6 Fusion Reference to Video

PixVerse

PixVerse V6 Fusion Reference to Video by PixVerse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pagePixVerse V6 Fusion Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

MiniMax H3 Text to Video

MiniMax

MiniMax H3 Text to Video by MiniMax can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

MiniMax H3 Image to Video

MiniMax

MiniMax H3 Image to Video by MiniMax can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

MiniMax H3 Reference to Video

MiniMax

MiniMax H3 Reference to Video by MiniMax can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageMiniMax H3 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

HappyHorse Text to Video

HappyHorse

HappyHorse Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHappyHorse Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

HappyHorse Image to Video

HappyHorse

HappyHorse Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHappyHorse Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

HappyHorse Reference to Video

HappyHorse

HappyHorse Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

HappyHorse Video Edit

HappyHorse

HappyHorse Video Edit by HappyHorse can transform an existing video without reshooting the source scene. Input: Text, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse Video Edit: Transform an existing video without reshooting the source scene.Input: Text, Video; output: Video.

HappyHorse 1.1 Image to Video

HappyHorse

HappyHorse 1.1 Image to Video by HappyHorse can animate a still image and turn it into a video. Input: Text, Images. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Image to Video: Animate a still image and turn it into a video.Input: Text, Images; output: Video.

HappyHorse 1.1 Text to Video

HappyHorse

HappyHorse 1.1 Text to Video by HappyHorse can create a video scene from a text brief. Input: Text. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Text to Video: Create a video scene from a text brief.Input: Text; output: Video.

HappyHorse 1.1 Reference to Video

HappyHorse

HappyHorse 1.1 Reference to Video by HappyHorse can build video from one or more visual references. Input: Text, Images, Video. Output: Video.

Video modelsExternal model, reference pageHappyHorse 1.1 Reference to Video: Build video from one or more visual references.Input: Text, Images, Video; output: Video.

Gemini Omni Video

Google

Gemini Omni Video by Google can create or transform video. Input: Text, Images, Video, Audio. Output: Video.

Video modelsExternal model, reference pageGemini Omni Video: Create or transform video.Input: Text, Images, Video, Audio; output: Video.

Gemini Omni Character

Google

Gemini Omni Character by Google can create or transform video. Input: Text, Images, Audio. Output: Video.

Video modelsExternal model, reference pageGemini Omni Character: Create or transform video.Input: Text, Images, Audio; output: Video.

OmniHuman 1.5

ByteDance

OmniHuman 1.5 by ByteDance can create or analyze a digital human and its presence in the frame. Input: Images, Audio, Video. Output: Video.

Video modelsExternal model, reference pageOmniHuman 1.5: Create or analyze a digital human and its presence in the frame.Input: Images, Audio, Video; output: Video.

Volcengine Video Lip Sync

ByteDance

Volcengine Video Lip Sync by ByteDance can synchronize lip movement in video with an audio track. Input: Video, Audio. Output: Video.

Video modelsExternal model, reference pageVolcengine Video Lip Sync: Synchronize lip movement in video with an audio track.Input: Video, Audio; output: Video.

Suno Music Video

Suno

Suno Music Video by Suno can create, transform, or continue musical material. Input: Audio, Images. Output: Video.

Video modelsExternal model, reference pageSuno Music Video: Create, transform, or continue musical material.Input: Audio, Images; output: Video.

Image models

A reviewed catalog of image generation and editing models; Neiron availability is marked separately.

All models and modes

GPT Image 2.5

OpenAI

Describe a scene or add a source photo and specify what to change. Check the available quality and compare the result with the source before publishing.

Image modelsAvailable in NeironCreate an image from a prompt.Edit an uploaded photo.

Nano Banana

Google

Start with a concrete job: a post cover, a portrait variation, or a product on a different background. In Nano Banana, describe a scene from scratch or upload a photo and request changes. When using several references, explain the role of each.

Image modelsAvailable in NeironDraw a scene from a description without a source photo.Produce a variation of an uploaded image following your brief.

Nano Banana Pro

Google

Choose the output size before making a poster or banner: Nano Banana Pro offers 1K, 2K, and 4K. Then describe the composition or add product photos. Check lettering, logos, and small details in the finished file, not just its preview.

Image modelsAvailable in NeironSelect 1K, 2K, or 4K before creating an image.Build a poster or banner from a composition brief.

Seedream 5 Pro

ByteDance

With Seedream 5 Pro, describe a new scene or request an edit using uploaded images. For example, take the subject from one photo and the palette from another. Neiron provides a prompt field, reference uploads, and a choice of 1K or 2K output.

Image modelsAvailable in NeironCompose an image from a written brief and visual references.Describe changes to an object or scene in plain language.

FLUX.2 [klein] 4B

Black Forest Labs

FLUX.2 [klein] 4B generates images from text and supports multi-reference editing. Black Forest Labs releases this four-billion-parameter model under Apache-2.0. It is distinct from klein 9B variants and FP8 quantizations.

Image modelsExternal model, reference pageGenerate a new image from a written brief.Edit a source image following an instruction.

Seedream 3.0 Text to Image

ByteDance

Seedream 3.0 Text to Image by ByteDance can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageSeedream 3.0 Text to Image: Create an image from a text description.Input: Text; output: Images.

Seedream 4.0 Text to Image

ByteDance

Seedream 4.0 Text to Image by ByteDance can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageSeedream 4.0 Text to Image: Create an image from a text description.Input: Text; output: Images.

Seedream 4.0 Edit

ByteDance

Seedream 4.0 Edit by ByteDance can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageSeedream 4.0 Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Seedream 4.5 Text to Image

ByteDance

Seedream 4.5 Text to Image by ByteDance can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageSeedream 4.5 Text to Image: Create an image from a text description.Input: Text; output: Images.

Seedream 4.5 Edit

ByteDance

Seedream 4.5 Edit by ByteDance can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageSeedream 4.5 Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Seedream 5.0 Lite Text to Image

ByteDance

Seedream 5.0 Lite Text to Image by ByteDance can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageSeedream 5.0 Lite Text to Image: Create an image from a text description.Input: Text; output: Images.

Seedream 5.0 Lite Image to Image

ByteDance

Seedream 5.0 Lite Image to Image by ByteDance can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageSeedream 5.0 Lite Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Seedream 5.0 Pro Text to Image

ByteDance

Seedream 5.0 Pro Text to Image by ByteDance can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageSeedream 5.0 Pro Text to Image: Create an image from a text description.Input: Text; output: Images.

Seedream 5.0 Pro Image to Image

ByteDance

Seedream 5.0 Pro Image to Image by ByteDance can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageSeedream 5.0 Pro Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Seedream 5.0 Pro Layer Decomposition

ByteDance

Seedream 5.0 Pro Layer Decomposition by ByteDance can decompose an image into editable visual layers. Input: Images. Output: Images.

Image modelsExternal model, reference pageSeedream 5.0 Pro Layer Decomposition: Decompose an image into editable visual layers.Input: Images; output: Images.

Z-Image

Tongyi Lab

Z-Image by Tongyi Lab can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageZ-Image: Create an image from a text description.Input: Text; output: Images.

Imagen 4 Fast

Google

Imagen 4 Fast by Google can perform a specialist operation on multimedia data. Input: Text. Output: Images.

Image modelsExternal model, reference pageImagen 4 Fast: Perform a specialist operation on multimedia data.Input: Text; output: Images.

Imagen 4 Ultra

Google

Imagen 4 Ultra by Google can perform a specialist operation on multimedia data. Input: Text. Output: Images.

Image modelsExternal model, reference pageImagen 4 Ultra: Perform a specialist operation on multimedia data.Input: Text; output: Images.

Imagen 4

Google

Imagen 4 by Google can perform a specialist operation on multimedia data. Input: Text. Output: Images.

Image modelsExternal model, reference pageImagen 4: Perform a specialist operation on multimedia data.Input: Text; output: Images.

Nano Banana Edit

Google

Nano Banana Edit by Google can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageNano Banana Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Nano Banana 2

Google

Nano Banana 2 by Google can perform a specialist operation on multimedia data. Input: Text. Output: Images.

Image modelsExternal model, reference pageNano Banana 2: Perform a specialist operation on multimedia data.Input: Text; output: Images.

Nano Banana 2 Lite

Google

Nano Banana 2 Lite by Google can perform a specialist operation on multimedia data. Input: Text. Output: Images.

Image modelsExternal model, reference pageNano Banana 2 Lite: Perform a specialist operation on multimedia data.Input: Text; output: Images.

FLUX.2 Pro Image to Image

Black Forest Labs

FLUX.2 Pro Image to Image by Black Forest Labs can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageFLUX.2 Pro Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

FLUX.2 Pro Text to Image

Black Forest Labs

FLUX.2 Pro Text to Image by Black Forest Labs can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageFLUX.2 Pro Text to Image: Create an image from a text description.Input: Text; output: Images.

FLUX.2 Image to Image

Black Forest Labs

FLUX.2 Image to Image by Black Forest Labs can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageFLUX.2 Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

FLUX.2 Text to Image

Black Forest Labs

FLUX.2 Text to Image by Black Forest Labs can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageFLUX.2 Text to Image: Create an image from a text description.Input: Text; output: Images.

Grok Imagine Text to Image

xAI

Grok Imagine Text to Image by xAI can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageGrok Imagine Text to Image: Create an image from a text description.Input: Text; output: Images.

Grok Imagine Image to Image

xAI

Grok Imagine Image to Image by xAI can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageGrok Imagine Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Grok Imagine Image 2.0 Text to Image

xAI

Grok Imagine Image 2.0 Text to Image by xAI can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageGrok Imagine Image 2.0 Text to Image: Create an image from a text description.Input: Text; output: Images.

Grok Imagine Image 2.0 Image to Image

xAI

Grok Imagine Image 2.0 Image to Image by xAI can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageGrok Imagine Image 2.0 Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

GPT Image 1.5 Text to Image

OpenAI

GPT Image 1.5 Text to Image by OpenAI can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageGPT Image 1.5 Text to Image: Create an image from a text description.Input: Text; output: Images.

GPT Image 1.5 Image to Image

OpenAI

GPT Image 1.5 Image to Image by OpenAI can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageGPT Image 1.5 Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

GPT Image 2 Text to Image

OpenAI

GPT Image 2 Text to Image by OpenAI can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageGPT Image 2 Text to Image: Create an image from a text description.Input: Text; output: Images.

GPT Image 2 Image to Image

OpenAI

GPT Image 2 Image to Image by OpenAI can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageGPT Image 2 Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Topaz Image Upscale

Topaz Labs

Topaz Image Upscale by Topaz Labs can increase the resolution and visual clarity of source media. Input: Images. Output: Images.

Image modelsExternal model, reference pageTopaz Image Upscale: Increase the resolution and visual clarity of source media.Input: Images; output: Images.

Recraft Remove Background

Recraft

Recraft Remove Background by Recraft can automatically separate a subject from its background. Input: Images. Output: Images.

Image modelsExternal model, reference pageRecraft Remove Background: Automatically separate a subject from its background.Input: Images; output: Images.

Recraft Crisp Upscale

Recraft

Recraft Crisp Upscale by Recraft can increase the resolution and visual clarity of source media. Input: Images. Output: Images.

Image modelsExternal model, reference pageRecraft Crisp Upscale: Increase the resolution and visual clarity of source media.Input: Images; output: Images.

Ideogram Character Edit

Ideogram

Ideogram Character Edit by Ideogram can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageIdeogram Character Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Ideogram Character Remix

Ideogram

Ideogram Character Remix by Ideogram can create a new image variation from a source visual. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageIdeogram Character Remix: Create a new image variation from a source visual.Input: Text, Images; output: Images.

Ideogram Character

Ideogram

Ideogram Character by Ideogram can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageIdeogram Character: Create an image from a text description.Input: Text; output: Images.

Ideogram V3 Text to Image

Ideogram

Ideogram V3 Text to Image by Ideogram can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageIdeogram V3 Text to Image: Create an image from a text description.Input: Text; output: Images.

Ideogram V3 Edit

Ideogram

Ideogram V3 Edit by Ideogram can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageIdeogram V3 Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Ideogram V3 Remix

Ideogram

Ideogram V3 Remix by Ideogram can create a new image variation from a source visual. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageIdeogram V3 Remix: Create a new image variation from a source visual.Input: Text, Images; output: Images.

Qwen Text to Image

Alibaba Cloud

Qwen Text to Image by Alibaba Cloud can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageQwen Text to Image: Create an image from a text description.Input: Text; output: Images.

Qwen Image to Image

Alibaba Cloud

Qwen Image to Image by Alibaba Cloud can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageQwen Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Qwen Image Edit

Alibaba Cloud

Qwen Image Edit by Alibaba Cloud can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageQwen Image Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Qwen2 Image Edit

Alibaba Cloud

Qwen2 Image Edit by Alibaba Cloud can apply requested changes to an existing image. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageQwen2 Image Edit: Apply requested changes to an existing image.Input: Text, Images; output: Images.

Qwen2 Text to Image

Alibaba Cloud

Qwen2 Text to Image by Alibaba Cloud can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageQwen2 Text to Image: Create an image from a text description.Input: Text; output: Images.

Qwen3 Pro Text to Image

Alibaba Cloud

Qwen3 Pro Text to Image by Alibaba Cloud can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageQwen3 Pro Text to Image: Create an image from a text description.Input: Text; output: Images.

Qwen3 Text to Image

Alibaba Cloud

Qwen3 Text to Image by Alibaba Cloud can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference pageQwen3 Text to Image: Create an image from a text description.Input: Text; output: Images.

Qwen3 Pro Image to Image

Alibaba Cloud

Qwen3 Pro Image to Image by Alibaba Cloud can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageQwen3 Pro Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

Qwen3 Image to Image

Alibaba Cloud

Qwen3 Image to Image by Alibaba Cloud can create a new image guided by a source frame. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageQwen3 Image to Image: Create a new image guided by a source frame.Input: Text, Images; output: Images.

4o Image API

OpenAI

4o Image API by OpenAI can create an image from a text description. Input: Text. Output: Images.

Image modelsExternal model, reference page4o Image API: Create an image from a text description.Input: Text; output: Images.

Flux Kontext

Black Forest Labs

Flux Kontext by Black Forest Labs can perform a specialist operation on multimedia data. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageFlux Kontext: Perform a specialist operation on multimedia data.Input: Text, Images; output: Images.

Wan 2.7 Image

Alibaba Cloud

Wan 2.7 Image by Alibaba Cloud can perform a specialist operation on multimedia data. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageWan 2.7 Image: Perform a specialist operation on multimedia data.Input: Text, Images; output: Images.

Wan 2.7 Image Pro

Alibaba Cloud

Wan 2.7 Image Pro by Alibaba Cloud can perform a specialist operation on multimedia data. Input: Text, Images. Output: Images.

Image modelsExternal model, reference pageWan 2.7 Image Pro: Perform a specialist operation on multimedia data.Input: Text, Images; output: Images.

Text models

A reviewed catalog of market language models for chat, translation, and reasoning; Neiron availability is marked separately.

All models and modes

GPT-6 Astra

OpenAI

Astra is useful when a task has several constraints and a short answer is not enough: compare options, work through material, or inspect a solution. State the criteria, inputs, and desired format first. For current facts, request links and dates, then check important conclusions against the primary sources.

Text modelsAvailable in NeironCompare options against stated criteria and flag missing information.Work through text, a table, or code and turn the findings into an action plan.

GPT-5.6 Luna

OpenAI

Start with Luna for an email from notes, a short summary, or an explanation of a small code fragment. Specify the desired length and audience. In Neiron, search and reasoning are separate controls; selecting the model does not enable them.

Text modelsAvailable in NeironTurn supplied points into an email or working draft.Shorten text while retaining the facts and numbers you specify.

GPT-5.6 Terra

OpenAI

Give Terra several constraints and ask for a structured decision: compare proposals, examine requirements, or outline a document. Separate facts, assumptions, and selection criteria upfront so the answer can be checked against your inputs.

Text modelsAvailable in NeironCompare options against specified criteria.Turn notes into a document structure or plan.

GPT-5.6 Sol

OpenAI

Use Sol for tasks with several dependencies: finding contradictions in a plan, investigating a failure, or comparing architectures. Supply the constraints and a correctness criterion. A detailed explanation alone does not establish that a conclusion is correct.

Text modelsAvailable in NeironReview an argument and identify unsupported steps.Propose failure hypotheses and checks that distinguish them.

Gemini 3.8 Flash

Google

Ask Flash to examine a screenshot, shorten notes, or explain a term with an example. For an image, specify whether to inspect labels, an interface problem, or differences between versions. Paste small text separately when accuracy matters.

Text modelsAvailable in NeironExamine visible elements in an uploaded image.Prepare a draft or shortened version of text.

Claude Sonnet 5

Anthropic

Give Sonnet an editing brief with several constraints: retain facts, remove repetition, and preserve the author’s position. For code, add expected behavior and error examples. Neiron runs this model without a separate thinking mode.

Text modelsAvailable in NeironEdit a document against specified rules and tone.Build a coherent draft from scattered notes.

Grok 4.6

xAI

Use text-based Grok to investigate a question, examine arguments, and prepare an answer with search context. Specify a period and a concrete question rather than asking for everything. Grok Imagine video is a separate mode with different results and ratings.

Text modelsAvailable in NeironPrepare a text analysis with search context.Compare arguments and identify questions about the sources.

Gemini 3.1 Pro

Google

When an explanation is in text but details are in an image, give Gemini both and ask a specific question. For example, compare page requirements with a screenshot. Request a clear distinction between directly visible evidence and interpretation.

Text modelsAvailable in NeironCompare a written description with an image.Analyze supplied materials and identify inconsistencies.

DeepSeek V4.1 Flash

DeepSeek

DeepSeek V4.1 Flash helps explain complex topics, review code, and compare options. Select it and continue on the web or in Telegram. Saved Flash chats work as before; V4 Pro remains a separate choice.

Text modelsAvailable in NeironSeparate a problem into inputs, constraints, and questions.Explain the logic of a code fragment.

DeepSeek V4 PRO

DeepSeek

Pro is a separate DeepSeek V4 variant for multi-step analysis. Use it to compare algorithms, investigate failures, or challenge architectural assumptions. Start with constraints and request a testable conclusion rather than merely a longer explanation.

Text modelsAvailable in NeironReview a solution for edge cases.Compare algorithms against stated constraints.

GigaChat

Sber

GigaChat is Sber's language model. It is not available in Neiron; GPT-5.6 Terra covers similar tasks.

Text modelsNot available in NeironRussian-language dialogueWriting and editing

Amazon Nova Premier

Amazon

Amazon Nova Premier is the senior model of the Amazon Nova family for analyzing text, images, and video inside the AWS stack; it returns text.

Text modelsExternal model, reference pageAmazon Nova Premier: Perform a specialist operation on multimedia data.Input: Text, Images, Video; output: Text.

Claude Opus 4.8

Anthropic

Claude Opus 4.8 accepts text and images and returns text. Anthropic still lists the model as available, but marks it as legacy.

Text modelsExternal model, reference pageClaude Opus 4.8: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images; output: Text.

Claude Opus 5

Anthropic

Claude Opus 5 is Anthropic's advanced model for long-horizon agentic work and coding; its official overview lists up to a 1M-token context and 128K-token output.

Text modelsExternal model, reference pageClaude Opus 5: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Command A

Cohere

Command A is Cohere's model for tool use, agentic scenarios, and multilingual text tasks.

Text modelsExternal model, reference pageCommand A: Work with text, instructions, and conversational requests.Input: Text; output: Text.

Llama 4 Maverick

Meta

Llama 4 Maverick is Meta's open Mixture-of-Experts model for general text tasks and image understanding, with weights available for study and deployment.

Text modelsExternal model, reference pageLlama 4 Maverick: Work with text, instructions, and conversational requests.Input: Text, Images; output: Text.

Llama 4 Scout

Meta

Llama 4 Scout is Meta's compact open model from the Llama 4 family for fast text tasks and image analysis with lighter resource needs.

Text modelsExternal model, reference pageLlama 4 Scout: Work with text, instructions, and conversational requests.Input: Text, Images; output: Text.

Mistral Large 3

Mistral AI

Mistral Large 3 is a multimodal Mistral AI model that accepts text and images and returns text.

Text modelsExternal model, reference pageMistral Large 3: Perform a specialist operation on multimedia data.Input: Text, Images; output: Text.

Kimi K2 Thinking

Moonshot AI

Kimi K2 Thinking is an open reasoning model from Moonshot AI for competition-grade problems, long reasoning chains, and large contexts.

Text modelsExternal model, reference pageKimi K2 Thinking: Perform multi-step analysis and produce a reasoned answer.Input: Text; output: Text.

Kimi-K3

Moonshot AI

Kimi K3 is an open-weight multimodal model from Moonshot AI. It accepts text, images, and video. The developer specifies a million-token context, 2.8 trillion total parameters, and 104 billion active parameters. The described agent workflows require a separate tool environment.

Text modelsExternal model, reference pageAnalyze written and visual material together.Work with long technical briefs and code.

Qwen3.8-2.4T-A95B

Qwen

Qwen3.8-2.4T-A95B provides the weights of a text-based MoE model with 2.4 trillion total and 95 billion active parameters. It targets writing, coding, and multi-step work. It is not identical to the Qwen3.8-Max API, whose description separately lists visual input and built-in tools.

Text modelsExternal model, reference pageAnalyze and generate written material.Analyze programming tasks with multiple constraints.

Qwen3.8-27B

Qwen

Qwen3.8-27B is a dense model with a vision encoder. It accepts text, images, and video and returns text. The model card specifies 27 billion parameters and a native 262,144-token context. Its weights can be deployed under Apache-2.0.

Text modelsExternal model, reference pageExplain diagrams, screenshots, and video in response to a question.Draft code and technical documents.

Qwen3.8-Flash-Next

Qwen

Qwen3.8-Flash-Next is a preview of a new open-weight Qwen architecture with visual input. The developer counts the language component, n-gram embeddings, and MTP separately: 125, 51, and 4 billion parameters. Its native context is 262,144 tokens; a larger window requires separate configuration.

Text modelsExternal model, reference pageAnswer text questions using an image.Work with long context within the chosen configuration.

GigaChat 2 Max

Sber

GigaChat 2 Max is Sber's senior GigaChat model for demanding text tasks, document analysis, and Russian-language work. Beyond text, it understands images and audio.

Text modelsExternal model, reference pageGigaChat 2 Max: Perform a specialist operation on multimedia data.Input: Text, Images, Audio; output: Text, Images.

T-pro IT 1.0

T-Tech

T-pro IT 1.0 provides open Russian-language weights from T-Tech for further fine-tuning and self-hosting. Its model card explicitly says it is not a ready-to-use conversational assistant.

Text modelsExternal model, reference pageT-pro IT 1.0: Work with text, instructions, and conversational requests.Input: Text; output: Text.

GLM-5.3

Z.ai

GLM-5.3 is a Z.ai text model for programming and multi-step work. It uses the GLM-5.2 base model with changes introduced through post-training. Coding-agent evaluation results depend on the model as well as its harness, tools, and execution budget.

Text modelsExternal model, reference pageAnalyze and write code from a written brief.Plan a sequence of technical actions.

GLM-5.3-Flash

Z.ai

GLM-5.3 Flash from Z.ai accepts text and images and produces text. The developer reports 320 billion parameters, with 18 billion active per token. Reasoning effort is a separate choice: low, high, or max. The Flash name does not establish response time in a particular service.

Text modelsExternal model, reference pageAnalyze an image together with a written question.Explain code and draft technical material.

Suno Lyrics

Suno

Suno Lyrics by Suno can create or revise song lyrics. Input: Text. Output: Text.

Text modelsExternal model, reference pageSuno Lyrics: Create or revise song lyrics.Input: Text; output: Text.

GPT-5.2

OpenAI

GPT-5.2 by OpenAI can perform multi-step analysis and produce a reasoned answer. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageGPT-5.2: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images; output: Text.

Claude Opus 4.7

Anthropic

Claude Opus 4.7 by Anthropic can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Opus 4.7: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Claude Fable 5

Anthropic

Claude Fable 5 by Anthropic can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Fable 5: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Claude Haiku 4.5

Anthropic

Claude Haiku 4.5 by Anthropic can work with text, instructions, and conversational requests. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Haiku 4.5: Work with text, instructions, and conversational requests.Input: Text, Images; output: Text.

Claude Opus 4.5

Anthropic

Claude Opus 4.5 by Anthropic can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Opus 4.5: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Claude Opus 4.6

Anthropic

Claude Opus 4.6 by Anthropic can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Opus 4.6: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Claude Sonnet 4.5

Anthropic

Claude Sonnet 4.5 by Anthropic can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageClaude Sonnet 4.5: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

GPT Codex

OpenAI

GPT Codex by OpenAI can handle coding, code analysis, and agentic development tasks. Input: Text, Images. Output: Text.

Text modelsExternal model, reference pageGPT Codex: Handle coding, code analysis, and agentic development tasks.Input: Text, Images; output: Text.

Gemini 2.5 Pro OpenAI-compatible

Google

Gemini 2.5 Pro OpenAI-compatible by Google can perform multi-step analysis and produce a reasoned answer. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 2.5 Pro OpenAI-compatible: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images, Audio, Video; output: Text.

Gemini 3 Pro OpenAI-compatible

Google

Gemini 3 Pro OpenAI-compatible by Google can perform multi-step analysis and produce a reasoned answer. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3 Pro OpenAI-compatible: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.1 Pro OpenAI-compatible

Google

Gemini 3.1 Pro OpenAI-compatible by Google can perform multi-step analysis and produce a reasoned answer. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.1 Pro OpenAI-compatible: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images, Audio, Video; output: Text.

Gemini 2.5 Flash OpenAI-compatible

Google

Gemini 2.5 Flash OpenAI-compatible by Google can work with text, instructions, and conversational requests. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 2.5 Flash OpenAI-compatible: Work with text, instructions, and conversational requests.Input: Text, Images, Audio, Video; output: Text.

Gemini 3 Flash OpenAI-compatible

Google

Gemini 3 Flash OpenAI-compatible by Google can perform multi-step analysis and produce a reasoned answer. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3 Flash OpenAI-compatible: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.5 Flash

Google

Gemini 3.5 Flash by Google can handle coding, code analysis, and agentic development tasks. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.5 Flash: Handle coding, code analysis, and agentic development tasks.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.5 Flash OpenAI-compatible

Google

Gemini 3.5 Flash OpenAI-compatible by Google can handle coding, code analysis, and agentic development tasks. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.5 Flash OpenAI-compatible: Handle coding, code analysis, and agentic development tasks.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.6 Flash OpenAI-compatible

Google

Gemini 3.6 Flash OpenAI-compatible by Google can handle coding, code analysis, and agentic development tasks. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.6 Flash OpenAI-compatible: Handle coding, code analysis, and agentic development tasks.Input: Text, Images, Audio, Video; output: Text.

Gemini 3 Flash v1beta

Google

Gemini 3 Flash v1beta by Google can perform multi-step analysis and produce a reasoned answer. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3 Flash v1beta: Perform multi-step analysis and produce a reasoned answer.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.7 Flash

Google

Gemini 3.7 Flash by Google can handle coding, code analysis, and agentic development tasks. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.7 Flash: Handle coding, code analysis, and agentic development tasks.Input: Text, Images, Audio, Video; output: Text.

Gemini 3.7 Flash OpenAI-compatible

Google

Gemini 3.7 Flash OpenAI-compatible by Google can handle coding, code analysis, and agentic development tasks. Input: Text, Images, Audio, Video. Output: Text.

Text modelsExternal model, reference pageGemini 3.7 Flash OpenAI-compatible: Handle coding, code analysis, and agentic development tasks.Input: Text, Images, Audio, Video; output: Text.

Audio models

AI models for speech recognition, synthesis, and audio processing with verifiable specifications and sources.

All models and modes

Gemini Omni Audio

Google

Gemini Omni Audio by Google can create or transform audio material. Input: Audio, Images. Output: Audio.

Audio modelsExternal model, reference pageGemini Omni Audio: Create or transform audio material.Input: Audio, Images; output: Audio.

ElevenLabs Audio Isolation

ElevenLabs

ElevenLabs Audio Isolation by ElevenLabs can isolate useful voice or sound from a noisy recording. Input: Audio. Output: Audio.

Audio modelsExternal model, reference pageElevenLabs Audio Isolation: Isolate useful voice or sound from a noisy recording.Input: Audio; output: Audio.

ElevenLabs Text to Dialogue v3

ElevenLabs

ElevenLabs Text to Dialogue v3 by ElevenLabs can voice a dialogue with multiple speakers. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageElevenLabs Text to Dialogue v3: Voice a dialogue with multiple speakers.Input: Text; output: Audio.

ElevenLabs Text to Speech Multilingual v2

ElevenLabs

ElevenLabs Text to Speech Multilingual v2 by ElevenLabs can turn text into natural-sounding speech. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageElevenLabs Text to Speech Multilingual v2: Turn text into natural-sounding speech.Input: Text; output: Audio.

ElevenLabs Text to Speech Turbo 2.5

ElevenLabs

ElevenLabs Text to Speech Turbo 2.5 by ElevenLabs can turn text into natural-sounding speech. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageElevenLabs Text to Speech Turbo 2.5: Turn text into natural-sounding speech.Input: Text; output: Audio.

Gemini 3.1 Flash Text to Speech

Google

Gemini 3.1 Flash Text to Speech by Google can turn text into natural-sounding speech. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageGemini 3.1 Flash Text to Speech: Turn text into natural-sounding speech.Input: Text; output: Audio.

Gemini 2.5 Pro Text to Speech

Google

Gemini 2.5 Pro Text to Speech by Google can turn text into natural-sounding speech. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageGemini 2.5 Pro Text to Speech: Turn text into natural-sounding speech.Input: Text; output: Audio.

Suno Music Generation

Suno

Suno Music Generation by Suno can create, transform, or continue musical material. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Music Generation: Create, transform, or continue musical material.Input: Text; output: Audio.

Suno Music Extension

Suno

Suno Music Extension by Suno can extend an existing video or music segment. Input: Text, Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno Music Extension: Extend an existing video or music segment.Input: Text, Audio; output: Audio.

Suno Upload and Cover Audio

Suno

Suno Upload and Cover Audio by Suno can create, transform, or continue musical material. Input: Text, Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno Upload and Cover Audio: Create, transform, or continue musical material.Input: Text, Audio; output: Audio.

Suno Upload and Extend Audio

Suno

Suno Upload and Extend Audio by Suno can extend an existing video or music segment. Input: Text, Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno Upload and Extend Audio: Extend an existing video or music segment.Input: Text, Audio; output: Audio.

Suno Add Instrumental

Suno

Suno Add Instrumental by Suno can create, transform, or continue musical material. Input: Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno Add Instrumental: Create, transform, or continue musical material.Input: Audio; output: Audio.

Suno Add Vocals

Suno

Suno Add Vocals by Suno can separate a music recording into vocal and instrumental components. Input: Audio, Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Add Vocals: Separate a music recording into vocal and instrumental components.Input: Audio, Text; output: Audio.

Suno Replace Music Section

Suno

Suno Replace Music Section by Suno can transform an existing video without reshooting the source scene. Input: Audio, Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Replace Music Section: Transform an existing video without reshooting the source scene.Input: Audio, Text; output: Audio.

Suno Persona

Suno

Suno Persona by Suno can create, transform, or continue musical material. Input: Audio, Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Persona: Create, transform, or continue musical material.Input: Audio, Text; output: Audio.

Suno Mashup

Suno

Suno Mashup by Suno can create, transform, or continue musical material. Input: Audio, Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Mashup: Create, transform, or continue musical material.Input: Audio, Text; output: Audio.

Suno WAV Conversion

Suno

Suno WAV Conversion by Suno can create or transform audio material. Input: Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno WAV Conversion: Create or transform audio material.Input: Audio; output: Audio.

Suno Stem Separation

Suno

Suno Stem Separation by Suno can separate a music recording into vocal and instrumental components. Input: Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno Stem Separation: Separate a music recording into vocal and instrumental components.Input: Audio; output: Audio.

Suno MIDI Generation

Suno

Suno MIDI Generation by Suno can turn a musical idea into an editable MIDI representation. Input: Audio. Output: Audio.

Audio modelsExternal model, reference pageSuno MIDI Generation: Turn a musical idea into an editable MIDI representation.Input: Audio; output: Audio.

Suno Sounds

Suno

Suno Sounds by Suno can create, transform, or continue musical material. Input: Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Sounds: Create, transform, or continue musical material.Input: Text; output: Audio.

Suno Custom Voice

Suno

Suno Custom Voice by Suno can create or transform audio material. Input: Audio, Text. Output: Audio.

Audio modelsExternal model, reference pageSuno Custom Voice: Create or transform audio material.Input: Audio, Text; output: Audio.

Other models

Reviewed information pages for specialist AI models.

All models and modes

Choose an AI model for your task

Start with the result and source material. Category pages compare Neiron settings; model pages provide prompts, limitations, and sources. There is no need to infer capabilities from Pro, Flash, or a version number.

Reading benchmarks

Check the task, exact variant, settings, and date. An Arena preference score is not an accuracy percentage or a test of our service. Missing data is not replaced with invented ratings.

Market catalog and Neiron availability

External cards help you research a model but do not add it to the product. Geography means the developer's country, not data residency or access from your location.

Compare plans

AI model FAQ

Which AI models are available in Neiron?
20 models and modes are currently available in Neiron for different tasks. Current availability and restrictions are shown on each card. An external reference card describes the market and does not itself mean the model is included with a Neiron plan.
How do I compare AI models and choose one?
Give models the same real task and the same inputs. Compare accuracy, format adherence, editability, speed, and allowance usage, then choose from your own result rather than the model name alone.
Do I need a separate subscription for each AI model?
No. Neiron's main plans cover models marked as available in the service. Text, image, and video allowances differ, so compare plans and the usage cost of the mode you need before paying.