‹ BackHN Continuity

Thread

GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price

1066 points · 953 comments · crorella

  1. the_duke · · focus · HN ↗
    The GPT 6 release was ... not great.

    Sol 6 was so bad that I switched over to Opus 5.5 exclusively.

    Huge regression compared to Sol 5.6, often doing really dumb things. Same for Luna.

    Even Astra is very unreliable for coding. Brilliant for vision, sometimes just great, but it also often does very stupid things.

    I'm a bit sour on OpenAI right now and skeptical that 6.1 will be much different.

    (Note: this is after preferring and shilling Codex/OpenAI models for the last half year)

    1. jstummbillig · · focus · HN ↗
      Eh. What? Is this common sentiment?

      I mean Opus 5.5 is absolutely fantastic, unreasonably and unexpectedly so, but Astra was great and as far as I can tell SOTA until, when was it, 3 days ago, no?

      (Sol 6 idk, have not used it much for coding really. Seemed to work just fine when Astra used it in Codex as subagents.)

      1. the_duke · · focus · HN ↗
        On r/codex the sentiment seems to be quite wide-spread.
        1. phoghed · · focus · HN ↗
          <a href="https:&#x2F;&#x2F;marginlab.ai&#x2F;trackers&#x2F;codex&#x2F;" rel="nofollow">https:&#x2F;&#x2F;marginlab.ai&#x2F;trackers&#x2F;codex&#x2F;

          Codex itself seems to have a regression. You can see clearly the token use changing significantly coincides with a score drop

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.