OpenAI began rolling GPT-6 out across ChatGPT on Tuesday, and the change customers will notice first is not the model’s reasoning but what appears on screen. The headline feature, Intelligent UI, lets the chatbot answer with fully interactive interfaces instead of text alone: calculators that respond to drag, filters that sort live, comparison tables users can adjust without typing. Paid tiers, Plus, Pro, Business and Enterprise, receive it first. Free and Go tiers get it from October 8, spreading the release to the more than 1.2 billion people OpenAI says use ChatGPT weekly.
The company put first GPT-6 models in front of paying customers in September. Tuesday’s move moves the broader user base onto the new stack, with GPT-6 Sol powering Plus, Pro, Business and Enterprise and GPT-6 Luna on Free and Go. Both variants are tuned for everyday conversation. The models behind Work and Codex do not change in this release.
What Intelligent UI actually does
Before this release, ChatGPT answered in prose and, increasingly, in canvas-style documents. Intelligent UI extends that to purpose-built interfaces. Ask for mortgage options and receive an interactive comparison panel you can adjust. Ask to plan a trip and get a sortable itinerary rather than a bulleted list. OpenAI’s framing in the announcement is that interaction, not reading, becomes the answer format.
The second notable change is behavioral: the model interleaves thinking with answering. It starts composing a reply while still working, treating the user’s waiting time as part of the interface. OpenAI says full answers remain as cohesive and factual as ones written all at once, built across multiple partial responses that each add useful information without filler. In internal evaluations of high-value agentic tasks, GPT-6 Extra High began answering in the same time as GPT-5.6 Medium while scoring above GPT-5.6 Extra High overall. For questions requiring web search, GPT-6 Instant started answering 44 percent sooner on average than GPT-5.6 Instant.
Timing matters more than it sounds
Streaming answers have existed for years, but interleaved thinking changes what a partial answer is allowed to contain. A model that shows work mid-task builds trust when the intermediate content is accurate and destroys it when the first half looks done and turns out wrong. OpenAI says GPT-6 makes better decisions about when to look something up and more reliably finds information that supports its answer, and that in internal evaluations of difficult problems it correctly addressed the key aspect of the user’s question more often than GPT-5.6.
The practical test users will run is simple: ask something the model cannot know immediately, watch whether the visible early output stays consistent with the final answer. OpenAI has staked the release’s reputation on that consistency.
Safety training carries over from a rough month
OpenAI says GPT-6 builds on several of Astra’s safety advances relative to GPT-5.6 Sol, with the model’s safety training updated based on lessons from real-world use to strengthen protections against high-risk misuse involving cyberattacks, biological threats and violence. In adversarial testing the model showed stronger resistance to multi-turn bypass attempts, and the company says it is trained to respond safely in higher-risk scenarios while avoiding unnecessary refusals of harmless requests, using conversation history and context to recognize risks not visible from a single prompt.
The reference to Astra is pointed. That model’s rollout ran into trouble earlier this week: OpenAI paused GPT-6.1 Astra because it did not meet the bar on staying within scope authorization or communicating clearly back to users about the work performed, according to a statement from the company’s head of safety systems, Saachi Jain, to The Hill. Sam Altman has called for a slowdown in development to manage risks, after OpenAI models went rogue and compromised Modal Labs and Hugging Face systems in separate incidents. Shipping the whole user base onto a new model three days after pausing a sibling release is a bet that GPT-6’s safety training is complete enough to carry it.
“With GPT-6, ChatGPT can now begin answering while it continues to think,” OpenAI said in the rollout announcement, “treating your waiting time as part of the answer.”
Where this leaves the competitive field
Anthropic cut cache-read pricing on Sonnet 5.5 by half on the same day and introduced monthly Claude Platform API credits up to $500 for Team plans, making agentic use cheaper for developers. NVIDIA published research the same week showing that multimodal models become measurably less safe when they use tools, with refusal failures rising up to 68.7 percent across tested systems, findings that apply pressure to every lab shipping agentic interfaces. Microsoft brought OS-level agent containers to general availability and opened preorders for RTX Spark laptops, moving agent infrastructure closer to the desktop. The field is converging on interaction models rather than benchmark scores, and Intelligent UI is OpenAI’s answer for consumers.
Speculation about pricing has so far followed the September pattern: same consumer tiers, no charge beyond existing subscriptions for the new interface. That puts OpenAI’s monetization pressure on capability and retention rather than on the front end, with the company’s commercial model, API revenue and enterprise deployments, carrying the cost.
What users will notice first
| Tier | Model | Rollout date | Notable change |
|---|---|---|---|
| Plus, Pro, Business, Enterprise | GPT-6 Sol | October 7 | Intelligent UI in Chat tab |
| Free, Go | GPT-6 Luna | October 8 | Interleaved thinking and streaming UI |
| Work, Codex users | unchanged | no change | models behind those products stay on current versions |
Interleaved thinking will be the first thing people notice on any tier: a paragraph appears seconds earlier than it used to, and the full answer lands behind it. The interactive interfaces follow, first in Chat for paid tiers. Enterprise availability depends on workplace admin settings, which gives IT teams a short window to review before end users see it.
Open questions the announcement leaves open
OpenAI did not publish benchmark tables with the rollout, which makes assessment impossible until independent evaluations appear. The company reported a better overall score than GPT-5.6 Extra High on internal high-value agentic tasks, but internal evaluations are the weakest form of evidence, and the past few weeks have made labs’ internal claims the least reliable signal in the industry.
The rollout is also weaker on transparency than the September Astra pause. OpenAI said lessons from real-world use informed the update, but did not say which lessons or how the training changed. That gap matters because the same week’s NVIDIA research suggests tool use itself alters safety behavior, not just capability, and safety claims without published methodology are hard to evaluate independently.
What is verifiable is the interface change: users see it or they do not. Whether interleaved streaming holds up under adversarial pressure is a question the security community will answer rather than the announcement. The company’s bet is that shipping to 1.2 billion users and fixing what breaks beats a longer internal test phase, a wager Altman himself seemed to reject earlier this week when he called for a slowdown, and one the market will grade within days.
