Skip to content
Trending storyPast story

OpenAI launches GPT-Live, a full-duplex voice model for simultaneous listening and speaking

4 reports4 sourcesupdated Jul 23, 2026

What happened

From the coverage

OpenAI 推出了新一代语音模型 GPT-Live,最大的变化是“全双工”——它能一边听你说话一边回话,不用等你闭嘴再开口。聊天时它会用“嗯”、“对”这类语气词表示在听,也能在你思考时安静等着。遇到需要上网搜资料或复杂推理的问题,GPT-Live 会把活儿派给后台的 GPT-5.5 去干,自己继续跟你聊,等结果回来再接上话。这次发了两个版本:GPT-...

From AI HOT 精选

Coverage

Follow the reports to see the story from different sides.

Sep 9
  1. Product Hunt · AI
    ChatGPT Images 2.5: Sharper visuals, faster flow, better creative control

    OpenAI launched ChatGPT Images 2.5 on Product Hunt, promising sharper visuals, faster generation, and better creative control. The post doesn't disclose technical details or benchmarks—just the tagline. Worth a test if you use ChatGPT for images, but take the hype with a grain of salt until hands-on reviews appear.

Jul 9
  1. Computing Life · Share · YagePick
    GPT-Live separates voice interaction from heavy reasoning—that's the real shift

    OpenAI launched GPT-Live on July 8, adding full-duplex and a delegation architecture to ChatGPT voice. Full-duplex lets you interrupt and talk while it works, but the principle isn't new—Moshi, Gemini Live, and ByteDance's Seeduplex all did it. The real change is delegation: the voice layer handles conversation while GPT-5.5 runs search, reasoning, and computation in parallel in the background, returning results as they arrive. This breaks the latency paradox where faster meant dumber. The voice model itself is limited—the System Card confirms it has no standalone tool access or code execution. No API yet; developers can only sign up for a waitlist. Realtime API remains the production workhorse at $64/M tokens for audio output. The post doesn't spell out whether custom tools can be plugged into the delegation layer or how much control developers will get over the black box.

  2. TechCrunch · AIPick
    OpenAI launches GPT-Live-1 voice models that can speak and listen at the same time

    OpenAI released GPT-Live-1 and GPT-Live-1 mini, full-duplex voice models that let users interrupt naturally and enable live translation. Paid ChatGPT users get mini by default; higher tiers can access the larger GPT-Live-1. The new models skip the old speech-to-text-to-speech pipeline and can call GPT-5.5 for search, reasoning, or agent tasks mid-conversation. Demos showed the model staying silent until summoned and presenting info visually. The post doesn't disclose pricing changes or a rollout timeline.

Jul 8
  1. AI HOT (Curated Pool)Pick
    OpenAI launches GPT-Live, a full-duplex voice model that listens and speaks at once

    OpenAI launched GPT-Live, a full-duplex voice model that can listen and speak simultaneously, rolling out to ChatGPT users today. It handles backchannels like 'mhmm,' pauses naturally, and delegates search or reasoning tasks to GPT-5.5 in the background while keeping the conversation going. Two versions are live: GPT-Live-1 and GPT-Live-1 mini. In 5–10 minute head-to-head tests, users strongly preferred GPT-Live over Advanced Voice Mode; it also scored higher on GPQA science reasoning and BrowseComp web search evals. API availability is not yet announced—developers can sign up for notifications.