\

Improving GPT-5.6 Sol in ChatGPT—and expanding access for free users

80 points - today at 5:02 PM

Source
  • sunaookami

    today at 9:14 PM

    This is actually a downgrade for free users since currently it uses GPT-5.5 for a few messages before it drops you down to GPT-5.5-mini. Now it always uses a model worse than Mini (Luna is nano-equivalent, "It roughly corresponds to the nano model tier used in earlier GPT-5 families." https://developers.openai.com/api/docs/models/gpt-5.6-luna ). I guess it's a bit better with Thinking though. They should use Terra for a few messages first before dropping down to Luna. And image inputs are still limited.

    • colingauvin

      today at 7:16 PM

      They must really be feeling the commoditization pressure. I'm not sure what the way out of this is, ChatGPT and Claude are still good products, but they are not necessarily premium products anymore.

      I expect a few things to happen in the next year:

      1) Exclusive MCP server deals/API integrations

      2) Significant switch to B2B marketing, even moreso than we've seen before, with API interfaces being paid and chat-client interfaces becoming more and more free, perhaps just with limits more on integrations or data visualization/analysis

      3) US restrictions on B2B contracts with non-US hosted models that do any sort of contracting with the government

      Obviously there's a bunch of stuff I'm not foreseeing. But it really does feel like the bottom of the market is collapsing into free. I assume OpenAI and Anthropic think their next generation of models will restore their halo tier status and that the cash burn is justified to just get there, but this has to really mess up IPO plans.

        • redox99

          today at 7:51 PM

          I'm not sure I agree

          1) Back then, even as a free user you'd be able to use the strongest model (even if with tight limits). Now, you need to pay to use Sol, and you need to pay to use Opus or Fable. It does seem fairly premium in that sense. Idk about 5.6 Luna, but the previous Instant was really bad, even for very casual users. It would hallucinate non stop.

          2) When $100 and $200 per month plans launched, they were received as outrageous even here. Nowadays they are pretty common among power users.

            • user43928

              today at 8:49 PM

              I also scoffed at the ChatGPT $200 Pro plan back then.

              Back when coding for me still meant copy-paste from the web version, it was only worth the $20/month for me.

              They only added the $100 Pro plan in April during GPT 5.4 times.

              Today I happily pay $400/month for Codex and Claude Code.

              • colingauvin

                today at 7:55 PM

                >1) Back then, even as a free user you'd be able to use the strongest model (even if with tight limits). Now, you need to pay to use Sol, and you need to pay to use Opus or Fable. It does seem fairly premium in that sense. Idk about 5.6 Luna, but the previous Instant was really bad, even for very casual users. It would hallucinate non stop.

                This is kind of what I'm saying though. Bottom has fallen out, differentiation is just can you be much more premium than the competition. Currently that remains unanswered.

                EDIT: I'm basing this off the assumption that for chat, premium is not a point of differentiation at all. For coding/analysis, it is.

                  • davidguetta

                    today at 8:37 PM

                    especially when 99% tasks don't need fable, and certainly not the next model

        • ilaksh

          today at 7:49 PM

          > Our mission is to ensure that artificial general intelligence benefits all of humanity. We’re introducing updates to ChatGPT that improve everyday conversations while expanding access for Free users.

          This clearly implies that they believe ChatGPT models are AGI and are now willing to say it out loud.

          Which I think is a fair interpretation of the term. They are general purpose intelligence in that you can get help from them about almost anything. They are not like narrow single purpose AI models.

          I don't think we need that term to mean "can completely emulate a human" or "can do every task any human on earth can do as well as them".

          It also needs to be differentiated from ASI with godlike powers many times greater than human.

            • kkoncevicius

              today at 8:09 PM

              If we could show the current models to someone like Alan Turing, I am sure he would conclude that we have AGI.

              • kubb

                today at 8:10 PM

                > This clearly implies that they believe ChatGPT models are AGI and are now willing to say it out loud.

                Well, the models are smart enough to point out why this is wrong.

            • heaney-555

              today at 7:55 PM

              Giving free ChatGPT users access to reasoning (the 'Think' toggle) will have a broader impact on the world than every new paid model and coding agent combined.

                • daemonologist

                  today at 8:28 PM

                  Going from 5.5 Instant (which was noticeably bad) to 5.6 Luna is a big jump as well. OpenAI is probably the most prominent among the general public - an advantage in some respects but they're giving away a lot of free inference and thus have to use a pretty small model to do it.

                    • johnsmith1840

                      today at 9:11 PM

                      I really don't see how they're going to be google here. The free tier is dominated by verticle integration and the platform that people use. Long term I imagine google wins the bottom of the market and I'd be suprised if they lost.

                  • gavinray

                    today at 9:03 PM

                    I'm not sure people fully comprehend the trickle-down pop-culture/zeitgeist effects that LLM's are having/are going to have on humanity.

                    Because everyone now outsources much of their thinking and researching to LLM's, our collective culture + brain is shaped in a cyclical manner by using them.

                    It's the mechanical homogenization of culture and groupthink.

                • firasd

                  today at 7:25 PM

                  I think it’s a misread to think the default ChatGPT model switching to GPT 5.6 Luna is some sort of desperation move. Keep in mind that Claude .ai never had this extreme stratification between the frontier models and the free tier (Sonnet is available to free users with rate limits).

                  So 5.6 Luna is just their next version of what they used to call 5.5 instant tier

                  And 5.x instant models were never much to write home about anyway so the default ChatGPT free model hasn’t been particularly distinctive since 4o

                  • ElijahLynn

                    today at 7:28 PM

                    I can't wait to never see a reasoning button ever again. Why do I have to reason about what reasoning level to use?

                      • skybrian

                        today at 7:42 PM

                        Some people want quick results. Some people want it to keep searching for a new math proof overnight without giving up, and they have money to burn.

                        It seems like giving it a time limit or a budget in dollars would be clearer, though?

                        Or, keep searching until I come back to the computer and ask about progress.

                          • pllbnk

                            today at 7:50 PM

                            The _Intelligence_ part of AGI should be able to guide the user through that without all the knobs.

                              • minimaxir

                                today at 7:57 PM

                                OpenAI tried auto-routing with the initial GPT-5 release and it was immediately clear why that was a bad idea.

                                • vanuatu

                                  today at 7:57 PM

                                  intelligence is not omniscience though

                          • redox99

                            today at 7:53 PM

                            Because the model can't read your mind and know if you want a quick answer, or an hour long deep dive.

                              • Jtarii

                                today at 8:04 PM

                                Then it should just ask the user what they want if its unclear from the context.

                                  • 2sk21

                                    today at 8:21 PM

                                    Exactly! I posted much the same comment in another thread and there were lots of huffy complaints that amounted to "you're prompting it wrong"

                                    • Sammi

                                      today at 8:35 PM

                                      That's what the reasoning slider is for!

                              • awakeasleep

                                today at 7:43 PM

                                Because your incentives are opposed to the provider’s incentives

                                • timpera

                                  today at 7:41 PM

                                  I think it's nice to be able to make the model reason for dozens of minutes when you want to go deep on a topic, even if the router thinks it's an easy question.

                              • simonw

                                today at 8:09 PM

                                Nothing in the ChatGPT model release notes yet: https://help.openai.com/en/articles/9624314-model-release-no...

                                (This is a subtle nudge at anyone from OpenAI who reads this to make sure they get updated.)

                                OpenAI have a model called "chat-latest" - I wonder if that's running this new model yet: https://developers.openai.com/api/docs/models/chat-latest

                                It's described as "points to the latest Instant model currently used in ChatGPT" - so presumably that's "GPT-5.6 Instant" in the app.

                                  • firasd

                                    today at 8:25 PM

                                    There’s no 5.6 instant I think — 5.6 Luna is gonna be the new instant tier model

                                      • simonw

                                        today at 8:52 PM

                                        No, Luna is the new free model.

                                        https://gist.github.com/simonw/aae4febd3c6f7bc5b7811857edb3c... has screenshots that still show "Instant" as an option for ChatGPT Chat... but not for ChatGPT Work.

                                          • firasd

                                            today at 9:03 PM

                                            Hmm

                                            It’s hard to understand this .. like sure we can select instant but is there an actual model called 5.6 instant? Like is it on LM Arena and OpenRouter or available via API etc

                                            5.5 instant is definitely A Thing it’s even name checked in this OAI post

                                • tosh

                                  today at 7:08 PM

                                  free unlimited luna is a pretty badass move

                                  luna is very good

                                    • timpera

                                      today at 7:54 PM

                                      Any improvement to the ChatGPT free plan is really nice. It's easy to forget that most people have never used a SOTA model, and only think of their experience with GPT-4o or Google Search's AI Overviews when asked about AI.

                                      • skybrian

                                        today at 7:28 PM

                                        It's also pretty cheap if you're paying for it (for coding). But I don't quite trust it for anything complicated, so I often use Terra.

                                          • redox99

                                            today at 8:06 PM

                                            Terra is awful and not cheap enough to make up for it. I'd suggest using luna or sol.

                                              • stuartq

                                                today at 9:14 PM

                                                Terra is far from awful. I've not used Luna enough to judge whether it's significantly better than Luna, but at least via GitHub Copilot, Terra is leaps ahead of Sonnet 5.

                                                • Sammi

                                                  today at 8:36 PM

                                                  So Luna is more awful but it's OK because it's cheaper?

                                          • causal

                                            today at 7:39 PM

                                            I guess I should give it another shot because I had pretty bad experience with Luna when it first came out. Stuff Sonnet knew better.

                                            • msq22

                                              today at 7:54 PM

                                              Were free users subject to some limits? I've never run into any limits.

                                                • timpera

                                                  today at 7:57 PM

                                                  It used to be 10 messages every 5 hours using GPT-5, then unlimited 4o-mini. More recently, it went down to ~5 messages per day on 5.5 Instant, then unlimited on 5.5 mini.

                                          • kingstnap

                                            today at 7:31 PM

                                            Its always fun to try to read between the lines here to speculate why they are doing this.

                                            Maybe Luna efficiency gain was actually significant enough that putting all the free users and giving them super generous limits makes sense.

                                            They might be doing this to improve the messaging of AI among causal users since right now there is a huge amount of datacenter backlash in the US due to AI grievances.

                                            Maybe they have too much excess capacity or they really want to juice token numbers and market share on their dashboards for marketing.

                                            I also wonder if being given access to an actually a decent model like luna with actual thinking budget instead of brainless "instant" modes will start to make causal users understand the real capabilities of these models.

                                              • planb

                                                today at 9:06 PM

                                                Let’s speculate: they want luna to be their first model “on silicon” and need as much test data as possible before finalizing the design.

                                                • deanc

                                                  today at 7:39 PM

                                                  It's also going to be more training data for them.

                                                  • ToValueFunfetti

                                                    today at 7:44 PM

                                                    Is this new behavior from them? I haven't looked at their free chat offering in a minute, but I thought they always had them close behind paid tier, often with essentially equal products that made it weird for them to sell the paid tier for chat.

                                                      • mkozlows

                                                        today at 7:47 PM

                                                        No, free tier was total garbage with GPT-5.

                                                        • kingstnap

                                                          today at 7:58 PM

                                                          You can't even pick 5.6 luna on the chat app with a paid subscription. It just gives you sol (or use older models) with what seems like basically as much usage as you want. And sol is considerably smarter than luna.

                                                          All of this is of pretty minor importance though. You can't read as many tokens as a subcription can produce so more chat is not the value add nor super important.

                                                          I mean there are literally so many providers for free chat if you are willing to use several seperate apps.

                                                          The real value in these subs is using codex cli, much like the real point of anthropic subs is using claude code. Because agentic work actually does require a lot of tokens.

                                                      • simianwords

                                                        today at 8:00 PM

                                                        > Maybe Luna efficiency gain was actually significant enough that putting all the free users and giving them super generous limits makes sense.

                                                        Definitely this. The recent 80% discount was a reaction to Deepseek's update so that they still position near the frontier. My theory: Luna has always had a much higher efficiency. You do know that the model didn't get faster after the discount?

                                                    • OsamaJaber

                                                      today at 8:39 PM

                                                      Expanding free access is mostly an inference cost

                                                      Serving cheaply at that scale means routing, batching, and cache hits, not a better model :D

                                                      • Squarex

                                                        today at 8:57 PM

                                                        Why pay for ChatGPT Go then?

                                                        • ignoramous

                                                          today at 7:19 PM

                                                            Every week, 1 billion people turn to ChatGPT for everything from quick questions and web searches to planning, research, advice, and complex decisions.
                                                          
                                                          Guess, Google's AI Mode is chipping away at their consumers (I know I haven't used Chat in a long, long while for 'quick questions and web searches' after OpenAI did away with "think" which I always use). The money-minting office & coding market Anthropic has cornered is hyper-competitive at both the frontier & low-cost ends. OpenAI is reactive [0] and seems right up against it, despite the strength of its excellent models.

                                                          [0] Won't put it past OpenAI (and/or Google) to open weight larger models!

                                                            • skybrian

                                                              today at 7:31 PM

                                                              $20/month also gives you API access for coding, in any coding agent. I keep hitting the weekly limit but it's a good deal while it lasts.

                                                              • drivebyhooting

                                                                today at 7:40 PM

                                                                Google’s AI has been very glitchy for me lately. I used to reserve chatGPT for serious work and Gemini for daily personalized unimportant things. But now I switched completely to ChatGPT and resigned myself to their memories/personalization.

                                                                • porridgeraisin

                                                                  today at 7:38 PM

                                                                  Did they do away with think? I think now you have to do it with /think

                                                              • simianwords

                                                                today at 8:01 PM

                                                                Does 5.6 Sol finally have an "instant" form? Is that the change?

                                                                  • today at 8:06 PM

                                                                    • redox99

                                                                      today at 8:06 PM

                                                                      Yes

                                                                  • jauntywundrkind

                                                                    today at 7:13 PM

                                                                    At first I thought the Sol updates was perhaps trying to help with some complaints of Sol burning through tokens, complaints that have prompted some new data points on https://codex-resets.com/ .

                                                                    But seeing the graphic with the visual weather report: that makes me think that is not the goal at all. :)

                                                                      • laweijfmvo

                                                                        today at 7:21 PM

                                                                        What a bizarre example of a more direct answer. Any human would simply say “No, it doesn’t rain here in summer.”

                                                                        Even after identifying the 0% chance of rain, it still drags the conversation on and on and on

                                                                          • sunaookami

                                                                            today at 9:10 PM

                                                                            GPT models were RLHF'd to death, they will never give a final, direct answer. Every release since the GPT-4o catastrophy is like this, it's so tiresome. They need a complete reset before it can become actually usable again. Or not as maximum engagement seems to be their goal.

                                                                            • taikahessu

                                                                              today at 7:36 PM

                                                                              Would you like to know more? Seriously though, of course, it's tuned for maximum engagement, not maximum efficiency. I wonder how long this engagement dopamine circus can last... too long apparently.

                                                                      • aniceperson

                                                                        today at 7:53 PM

                                                                        a m-dash in the title? whoa things are degrading fast

                                                                        • porridgeraisin

                                                                          today at 7:40 PM

                                                                          I dont pay for a chatgpt subscription, but sometimes I did use the web app for throwaway questions. GPT 5.5 Instant or whatever it was that they had was absolutely horrendous. Never answered a question straight and was pedantic in a way even a redditor wouldn't be. So I dropped it and just opened my paid coding agent for everything. Grok.com is quite good now with grok 4.5 though and I find myself using that often. Hopefully luna will be similar.