Best Audio AI Models

8 models tracked · 60 recent news stories

🏆

Most-talked-about Audio right now

Ranked by mentions across 30+ AI sources in October 2026.

GP
GPT-Live-1

GPT-Live-1 is OpenAI's full-duplex voice model, which can listen and speak at the same time. OpenAI introduced it as GPT-Live on July 8, 2026, powering ChatGPT's upgraded voice mode, and opened it to developers in the API on September 10, 2026 at $0.05 per minute for the voice layer. The model handles the conversation itself — turn-taking and back-channel cues — and hands the reasoning to a separate model the developer chooses.

GE
Gemini 3.8 Live

Gemini 3.8 Live is Google's real-time voice model, released on September 15, 2026 together with a Gemini 3.8 Live Extended Thinking variant. Google calls them its most advanced live dialogue models: they run tool and API calls in the background while the conversation continues, understand what the camera sees and support 97 languages. Both are available in the Gemini API, Google AI Studio and Gemini Enterprise, and Gemini 3.8 Live now powers Search Live in the Google app.

LY
Lyria 3

Lyria 3 helps you express, explore, and experiment with high-fidelity music, using prompts to create tracks with natural flow from note to note.

MU
Mureka V9.5

Mureka V9.5 is the AI music generation model from Mureka, the music platform of China's Kunlun Tech, introduced in July 2026. Kunlun pitches it as making songs that sound less machine-made by using "reflective reasoning" and agentic song creation, and in August 2026 Mureka launched it internationally alongside a second music model called O3.

SU
Suno
SU
Suno v6

Suno v6 is the AI music generation model Suno released on September 9, 2026 — its first model built with support from the record industry, following partnerships with Warner Music and BMG. Suno says v6 was trained from the ground up on a new, licensed dataset rather than the music used for earlier versions, and it replaces the retiring v4.5 and v5 models as the company faces a wave of copyright lawsuits.

UD
Udio

Udio's text-to-music model generates complete songs — vocals, lyrics and backing instrumentation — from a natural-language prompt.

WH
Whisper

OpenAI's speech-to-text AI model

📰 Latest Audio Model News(60 stories)