← All news·2026-08-09·3 min read

ChatGPT GPT-Live: Conversations With No Lags or Cutoffs

ChatGPT replaced the three-step voice pipeline with a single multimodal model, GPT-Live. Full-duplex: it speaks and listens simultaneously, inserts natural backchannels, and offloads complex tasks to GPT-5.5 in the background.

aichatgptopenaivoice

OpenAI has launched GPT-Live — ChatGPT's new voice engine. The old setup worked in three steps: speech was converted to text, a language model processed the text, and the response was synthesized back into voice. Each handoff added latency and risked losing context. GPT-Live replaces that pipeline with a single multimodal model that works directly with audio.

Full-Duplex and Natural Backchannels

The full-duplex architecture lets the model listen and speak at the same time — without awkward pauses when turns switch. GPT-Live adds natural reactions: it inserts "uh-huh" and "yeah" while the user is still talking. Complex requests are offloaded to GPT-5.5 in the background, keeping the conversation from stalling.

Availability and Intelligence Levels

GPT-Live is available on iOS, Android, and in the browser. Paid users (Pro, Plus, Go) get the full GPT-Live-1 with three modes: Instant, Medium, and High. The free plan is limited to GPT-Live-1 mini — Instant mode only. Live translation is already included; screen sharing is not yet available.

Source: www.engadget.com

Free course

Stop reading about AI — start building with it

The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.

Start free →
EAEvgenii Arsentev

Author

Evgenii Arsentev

PhD · Chief Executive Officer, digital health