So 5.6 Luna is just their next version of what they used to call 5.5 instant tier
And 5.x instant models were never much to write home about anyway so the default ChatGPT free model hasn’t been particularly distinctive since 4o
I expect a few things to happen in the next year:
1) Exclusive MCP server deals/API integrations
2) Significant switch to B2B marketing, even moreso than we've seen before, with API interfaces being paid and chat-client interfaces becoming more and more free, perhaps just with limits more on integrations or data visualization/analysis
3) US restrictions on B2B contracts with non-US hosted models that do any sort of contracting with the government
Obviously there's a bunch of stuff I'm not foreseeing. But it really does feel like the bottom of the market is collapsing into free. I assume OpenAI and Anthropic think their next generation of models will restore their halo tier status and that the cash burn is justified to just get there, but this has to really mess up IPO plans.
luna is very good
Maybe Luna efficiency gain was actually significant enough that putting all the free users and giving them super generous limits makes sense.
They might be doing this to improve the messaging of AI among causal users since right now there is a huge amount of datacenter backlash in the US due to AI grievances.
Maybe they have too much excess capacity or they really want to juice token numbers and market share on their dashboards for marketing.
I also wonder if being given access to an actually a decent model like luna with actual thinking budget instead of brainless "instant" modes will start to make causal users understand the real capabilities of these models.
But seeing the graphic with the visual weather report: that makes me think that is not the goal at all. :)
Every week, 1 billion people turn to ChatGPT for everything from quick questions and web searches to planning, research, advice, and complex decisions.
Guess, Google's AI Mode is chipping away at their consumers (I know I haven't used Chat in a long, long while for 'quick questions and web searches' after OpenAI did away with "think" which I always use). The money-minting office & coding market Anthropic has cornered is hyper-competitive at both the frontier & low-cost ends. OpenAI is reactive [0] and seems right up against it, despite the strength of its excellent models.[0] Won't put it past OpenAI (and/or Google) to open weight larger models!
It seems like giving it a time limit or a budget in dollars would be clearer, though?
Or, keep searching until I come back to the computer and ask about progress.
1) Back then, even as a free user you'd be able to use the strongest model (even if with tight limits). Now, you need to pay to use Sol, and you need to pay to use Opus or Fable. It does seem fairly premium in that sense. Idk about 5.6 Luna, but the previous Instant was really bad, even for very casual users. It would hallucinate non stop.
2) When $100 and $200 per month plans launched, they were received as outrageous even here. Nowadays they are pretty common among power users.
Even after identifying the 0% chance of rain, it still drags the conversation on and on and on
Back when coding for me still meant copy-paste from the web version, it was only worth the $20/month for me.
They only added the $100 Pro plan in April during GPT 5.4 times.
Today I happily pay $400/month for Codex and Claude Code.
Definitely this. The recent 80% discount was a reaction to Deepseek's update so that they still position near the frontier. My theory: Luna has always had a much higher efficiency. You do know that the model didn't get faster after the discount?
This is kind of what I'm saying though. Bottom has fallen out, differentiation is just can you be much more premium than the competition. Currently that remains unanswered.
EDIT: I'm basing this off the assumption that for chat, premium is not a point of differentiation at all. For coding/analysis, it is.
All of this is of pretty minor importance though. You can't read as many tokens as a subcription can produce so more chat is not the value add nor super important.
I mean there are literally so many providers for free chat if you are willing to use several seperate apps.
The real value in these subs is using codex cli, much like the real point of anthropic subs is using claude code. Because agentic work actually does require a lot of tokens.
Our mission is to ensure that artificial general intelligence benefits all of humanity. We’re introducing updates to ChatGPT that improve everyday conversations while expanding access for Free users.
For Plus and Pro users, we’re updating GPT‑5.6 Sol in Chat to be more reliable with facts and provide more focused answers. A new slider lets you choose how much thought ChatGPT puts into each response.
For Free users, we're updating the default model to GPT‑5.6 Luna and expanding access with unlimited text chats. For questions that need more thought, a new Think button lets you access higher reasoning for harder questions.
Every week, 1 billion people turn to ChatGPT for everything from quick questions and web searches to planning, research, advice, and complex decisions. We’ve updated GPT‑5.6 Sol to better support that full range. It delivers more focused answers, adapts its level of detail to the question, avoids unnecessary formatting, and offers a helpful correction when simply agreeing wouldn’t be useful. For Plus and Pro users, the same model now powers both Instant responses and deeper reasoning, creating one consistent experience.
The updates to GPT‑5.6 Sol in ChatGPT are designed to give you more direct responses, use tighter formatting, and avoid extra detail when it does not help.
For a quick question, that means a direct answer with the context you need. For more involved work like multi-step planning, research or writing, it means a fuller response that keeps the main recommendation clear.
A useful answer needs to get the facts right. The new GPT‑5.6 Sol is designed to make fewer mistakes—especially when answers depend on dates, numbers, sources, rules, or assumptions—by better using the sources it finds to answer your question.
In an internal evaluation of financial, medical, and legal prompts requiring factual detail, responses containing at least one factual error were about 62% less common with GPT‑5.6 Luna and 68% less common with GPT‑5.6 Sol than with GPT‑5.5 Instant.
With this update, we’re also bringing ChatGPT’s Instant and Thinking experiences closer together, creating a more consistent tone and behavior across different kinds of conversations. When you move from Instant to higher effort, it should feel like the model is taking extra time for a more comprehensive answer—not like you’re switching to a different model with its own tone or style.
Plus and Pro users can use the new slider in ChatGPT on web, mobile, and desktop to choose how much thought ChatGPT puts into an answer. Keep it quick for everyday questions, or move the slider up for planning, research, writing, coding, or decisions that need more thought.
We’re expanding access to our latest models for free users with unlimited text chats using GPT‑5.6 Luna, plus a new Think button for harder questions.
For questions that require deeper reasoning, Free users can tap the new Think button to give GPT‑5.6 Luna more time to work through the answer.

Plus and Pro users can access the updated version of GPT‑5.6 Sol and the new slider in ChatGPT starting today.
GPT‑5.6 Luna will become the default model for Free and Go users this week. Starting next week, they’ll also have unlimited text chats and access to a new Think button for harder questions (subject to abuse guardrails). Limits will still apply for file uploads, images and other tools.
Because this version of GPT‑5.6 Sol is optimized for everyday chats, it will only be available in the Chat experience in ChatGPT. The version of GPT‑5.6 Sol that powers Work and Codex is not changing as part of this release.
You can find more detail on safety training and evaluations in our system card(opens in a new window), which outlines additional measures we’ve introduced to support users we believe are under 18. For these users, we trained the model to avoid romantic roleplay, age-restricted challenges, and presenting itself as a substitute for real-world relationships. In addition, we applied age-appropriate boundaries around sexual content, eating disorders and body-image risks, age-restricted goods, dangerous activities, and graphic violence. Finally, the model encourages connection with trusted people when a teen may need support. We reinforced this training with system-level protections and have added new evaluations for how our models perform for users under 18, and are continuing to improve model responses in this area.
This is a concrete step toward more abundant intelligence: making our latest models more widely available, improving the usefulness and reliability of the answers people get, and letting free users keep text chats going without a rate limit. Access shapes opportunity, and this update gives more people the ability to keep asking, develop an idea, and get help when they need it.