‹ BackHN Continuity

Thread

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

80 points · 102 comments · theanonymousone

  1. atkrista · · focus · HN ↗
    I would just LOVE to see all the behind-the-scenes shithousery both companies are employing to one-up the other in this, largely, 2-horse AGI race. Someone should make a mockumentary when all is said and done!
    1. nbardy · · focus · HN ↗
      I think it's weirdly just a choice of deciding to cut releases.

      We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

      1. mFixman · · focus · HN ↗
        Any strong enough model with weak enough safeguards can cause an AI Chernobyl event that will make people and governments against AI development and deployment, just like Chernobyl did for nuclear energy.
        1. ChromeUltron · · focus · HN ↗
          tell me "I drank the kool aid" without telling me you drank the kool aid.
          1. mFixman · · focus · HN ↗
            The US government and most large companies drank the kool aid, and they will be the ones blaming Big AI if things go very wrong.
            1. esseph · · focus · HN ↗
              I think there's going to be a brutal backlash against the entire technology sector.

              Gov and Corp will throw their hands up and explain how it's not their fault.

      2. iLoveOncall · · focus · HN ↗
        > We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

        You're just believing their own bullshit. There's no indication that this is true except from claims from people working at OpenAI.

        If they really had a much more powerful model, it would make absolutely no sense to sit on it.

        1. adamzenith · · focus · HN ↗
          You don't think having a more intelligent model they can use internally that others can't is an advantage?
          1. iLoveOncall · · focus · HN ↗
            No? The top of human engineers are much better than any model would be, so AI models really aren't a big advantage when you're trying to develop anything that is SOTA.
            1. ceejayoz · · focus · HN ↗
              Not every problem is best addressed by a top engineer.

              Plenty of the tasks that keep a company running can benefit from good-enough (and better than the competition).

              1. iLoveOncall · · focus · HN ↗
                Yes, and none of those tasks require even the current SOTA models.
            2. x187463 · · focus · HN ↗
              You can't fathom how 'top human engineers' could take advantage of exclusive access to a frontier model at ultrafast speeds and limitless token budgets to conduct research?
              1. iLoveOncall · · focus · HN ↗
                I have all that and LLMs invariably produce garbage, so, no.
          2. owebmaster · · focus · HN ↗
            If that's the case, do Anthropic have an even better one helping them? The Chinese labs too? Where this OpenAI "advantage" is taking them?
        2. howdareme9 · · focus · HN ↗
          its not done training, why would they release a model that hasn't finished training?

          besides, we know anthropic are sitting on models too

          1. iLoveOncall · · focus · HN ↗
            > its not done training, why would they release a model that hasn't finished training?

            Because clearly they have no problem with releasing newer versions of models even just a week apart.

            > besides, we know anthropic are sitting on models too

            It is from your crystal ball or from other bullshit you heard from Anthropic employees on Twitter?

            We all know Anthropic had Mythos and Fable, and they turned out to be completely normal models, entirely in line with the capability of their predecessors.

            All they do is lie, and you're believing their lies.

            1. meowface · · focus · HN ↗
              The poster is not claiming Bel is a secret AGI. Just that it exists and only exists internally at the moment.

              It's rumored to be over 10T parameters. When released it'll probably be very good at certain tasks, albeit slow and expensive and not necessarily "wiser". You don't have to make this a binary.

              Also, Mythos was in fact a significant step-up in several ways. It fits the trend line, but only because the trend line for LLMs is quite steep. Plus Astra is still in many ways less intelligent than Fable/Mythos despite being released much later.

            2. azan_ · · focus · HN ↗
              > its not done training, why would they release a model that hasn't finished training?

              > Because clearly they have no problem with releasing newer versions of models even just a week apart.

              You can see how it is pure non-sequitur, right?

              > We all know Anthropic had Mythos and Fable, and they turned out to be completely normal models, entirely in line with the capability of their predecessors.

              Fable was absolutely not in line with capabilities of other models when released. For cybersec work it was much, much better.

        3. meowface · · focus · HN ↗
          With all due respect, you do not have a single clue what you're talking about.
          1. owebmaster · · focus · HN ↗
            With the same respect, you don't either. Simping for openai don't make you part of their in group
            1. meowface · · focus · HN ↗
              I definitely do not know what I'm talking about, but that individual doesn't know what they're talking about even more than I don't know what I'm talking about.
        4. simonw · · focus · HN ↗
          It makes sense for them to sit on it until they've finished testing it. More powerful but also more likely to delete all your email by mistake = you shouldn't release it yet.
        5. 233mhz · · focus · HN ↗
          > If they really had a much more powerful model, it would make absolutely no sense to sit on it.

          Makes total sense if they don't have the compute and can't serve it in an economically viable way. Also it lets you build things no one else in the world can build as fast as you until it's released

          1. iLoveOncall · · focus · HN ↗
            > Also it lets you build things no one else in the world can build as fast as you until it's released

            "We can generate slop faster than anyone in the world" :evil_emoji:

          2. owebmaster · · focus · HN ↗
            This theory would be easy to prove IF openai was pushing high quality software. It's not.
        6. _davide_ · · focus · HN ↗
          > it would make absolutely no sense to sit on it.

          Yeah, it does, it might be misaligned, a snapshot of the going on training, bigger than they can serve publicly, not yet completed the full training pipeline.

          If you see the knowledge cutoff you can see that sol 5.6 finished the initial main (+stage edited) of the training pipeline on Feb 16, 2026 but it was publicly released on July 9.

          The opposite would be weird: if they do NOT have an unreleased in-house model that would be really odd.

          1. iLoveOncall · · focus · HN ↗
            > The opposite would be weird: if they do NOT have an unreleased in-house model that would be really odd.

            This isn't at all what I said. The original commenter mentioned a model MUCH stronger than Astra.

          2. nextaccountic · · focus · HN ↗
            > The opposite would be weird: if they do NOT have an unreleased in-house model that would be really odd.

            Indeed it would be really odd if OpenAI were actually open

      3. 233mhz · · focus · HN ↗
        > We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

        I mean, isn't it almost a guarantee that what we get is a gimped version of what they use internally? They probably already serve themselves next gen level models at 1k+ tps from cerebras machines hosted on perm while we get quantized astra/opus at 50tps on a good day

        1. arcfour · · focus · HN ↗
          Was it not essentially confirmed that all of the frontier labs have much stronger models internally that don't make sense to serve at scale yet, which they use for development and to train models that are able to be served at scale?
          1. Alifatisk · · focus · HN ↗
            Where was this confirmed?
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.