Audio models
AI models for speech recognition, synthesis, and audio processing with verifiable specifications and sources.
Gemini Omni Audio
Gemini Omni Audio is part of Google's Gemini Omni family. This card is for an audio workflow. It can create or transform audio material. Consider it when the desired result is sound, speech, or music. Input: Audio, Images; output: Audio.
ElevenLabs Audio Isolation
ElevenLabs
ElevenLabs Audio Isolation is part of ElevenLabs's ElevenLabs family. This card is for an audio workflow. It can isolate useful voice or sound from a noisy recording. Consider it when the recording needs cleanup before editing, transcription, or publishing. Input: Audio; output: Audio.
ElevenLabs Text to Dialogue v3
ElevenLabs
ElevenLabs Text to Dialogue v3 is part of ElevenLabs's ElevenLabs family. This card is for an audio workflow. It can voice a dialogue with multiple speakers. Consider it when text needs to become a conversation, scene, or podcast segment. Input: Text; output: Audio.
ElevenLabs Text to Speech Multilingual v2
ElevenLabs
ElevenLabs Text to Speech Multilingual v2 is part of ElevenLabs's ElevenLabs family. This card is for an audio workflow. It can turn text into natural-sounding speech. Consider it when you need narration, a voice-over, or a spoken interface. Input: Text; output: Audio.
ElevenLabs Text to Speech Turbo 2.5
ElevenLabs
ElevenLabs Text to Speech Turbo 2.5 is part of ElevenLabs's ElevenLabs family. This card is for an audio workflow. It can turn text into natural-sounding speech. Consider it when you need narration, a voice-over, or a spoken interface. Input: Text; output: Audio.
Gemini 3.1 Flash Text to Speech
Gemini 3.1 Flash Text to Speech is part of Google's Gemini TTS family. This card is for an audio workflow. It can turn text into natural-sounding speech. Consider it when you need narration, a voice-over, or a spoken interface. Input: Text; output: Audio.
Gemini 2.5 Pro Text to Speech
Gemini 2.5 Pro Text to Speech is part of Google's Gemini TTS family. This card is for an audio workflow. It can turn text into natural-sounding speech. Consider it when you need narration, a voice-over, or a spoken interface. Input: Text; output: Audio.
Suno Music Generation
Suno
Suno Music Generation is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Text; output: Audio.
Suno Music Extension
Suno
Suno Music Extension is part of Suno's Suno family. This card is for an audio workflow. It can extend an existing video or music segment. Consider it when a successful source needs a coherent continuation. Input: Text, Audio; output: Audio.
Suno Upload and Cover Audio
Suno
Suno Upload and Cover Audio is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Text, Audio; output: Audio.
Suno Upload and Extend Audio
Suno
Suno Upload and Extend Audio is part of Suno's Suno family. This card is for an audio workflow. It can extend an existing video or music segment. Consider it when a successful source needs a coherent continuation. Input: Text, Audio; output: Audio.
Suno Add Instrumental
Suno
Suno Add Instrumental is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Audio; output: Audio.
Suno Add Vocals
Suno
Suno Add Vocals is part of Suno's Suno family. This card is for an audio workflow. It can separate a music recording into vocal and instrumental components. Consider it when separate tracks are needed for remixing, analysis, or a new mix. Input: Audio, Text; output: Audio.
Suno Replace Music Section
Suno
Suno Replace Music Section is part of Suno's Suno family. This card is for an audio workflow. It can transform an existing video without reshooting the source scene. Consider it when you need to change the style, subject, character, or part of a clip. Input: Audio, Text; output: Audio.
Suno Persona
Suno
Suno Persona is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Audio, Text; output: Audio.
Suno Mashup
Suno
Suno Mashup is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Audio, Text; output: Audio.
Suno WAV Conversion
Suno
Suno WAV Conversion is part of Suno's Suno family. This card is for an audio workflow. It can create or transform audio material. Consider it when the desired result is sound, speech, or music. Input: Audio; output: Audio.
Suno Stem Separation
Suno
Suno Stem Separation is part of Suno's Suno family. This card is for an audio workflow. It can separate a music recording into vocal and instrumental components. Consider it when separate tracks are needed for remixing, analysis, or a new mix. Input: Audio; output: Audio.
Suno MIDI Generation
Suno
Suno MIDI Generation is part of Suno's Suno family. This card is for an audio workflow. It can turn a musical idea into an editable MIDI representation. Consider it when the material needs further work in a sequencer or arrangement. Input: Audio; output: Audio.
Suno Sounds
Suno
Suno Sounds is part of Suno's Suno family. This card is for an audio workflow. It can create, transform, or continue musical material. Consider it when the task involves composition, vocals, arrangement, or sound design. Input: Text; output: Audio.
Suno Custom Voice
Suno
Suno Custom Voice is part of Suno's Suno family. This card is for an audio workflow. It can create or transform audio material. Consider it when the desired result is sound, speech, or music. Input: Audio, Text; output: Audio.
How to read this category
This page groups models of one type, but they solve different jobs. Start from the use case, not the model name: it is the shortest path to a usable result without unnecessary reruns.
- check purpose, modalities, and developer against reviewed sources
- separate models available in Neiron from market information pages
Practical choice
- Availability in Neiron is confirmed only by the internal product catalog.
- Unverified benchmarks and specifications are not published.