OpenAI’s GPT‑Live‑1 powers the new ChatGPT Voice experience (web search, memory, and multimodal chat)
OpenAI’s latest ChatGPT Voice rollout is powered by GPT‑Live‑1 for paid users and GPT‑Live‑1 mini for Free users, enabling simultaneous listening and speaking and adding support for web search, memory, and mixed text+image conversations in Voice.
OpenAI published a short demo highlighting “improved intelligence” in GPT‑Live, positioning the model as able to keep a conversation going while helping with multiple tasks at once (e.g., checking flights, looking up weather, shaping an itinerary) (OpenAI on X).
For durable details, OpenAI’s ChatGPT release notes describe a new ChatGPT Voice experience powered by GPT‑Live‑1 for paid users and GPT‑Live‑1 mini for Free users (ChatGPT release notes).
What’s new (verified)
- Voice runs on GPT‑Live‑1 / GPT‑Live‑1 mini. OpenAI says ChatGPT Voice is now powered by GPT‑Live‑1 (paid) and GPT‑Live‑1 mini (Free) (ChatGPT release notes).
- Simultaneous listening + speaking. The release notes say both models “can listen and speak at the same time,” making interruptions and turn-taking feel more natural (ChatGPT release notes).
- Tools and multimodality inside Voice. OpenAI says GPT‑Live‑1 can use web search and memory, show visual results via supported widgets, and “work with text and images in the same conversation” (ChatGPT release notes).
Availability notes (verified)
OpenAI says GPT‑Live‑1 is rolling out across consumer plans (including Free) on chatgpt.com and the ChatGPT iOS and Android apps in supported regions (ChatGPT release notes).
It is not available in ChatGPT Business, Enterprise, or Edu workspaces at launch, and it “does not support video or screen sharing at this time” (ChatGPT release notes).
Why it matters
If you use ChatGPT Voice as a real-time copilot, the key shift is that Voice is no longer only about speaking an answer: OpenAI is explicitly framing GPT‑Live as a model that can keep the conversational thread while you juggle multiple tasks (OpenAI on X).
That makes Voice more useful for planning and decision flows where you need context to persist across turns — especially when you want the assistant to combine live lookups (web search) with personal context (memory) while still letting you interrupt and steer the conversation (ChatGPT release notes).