‹ BackHN Continuity

Thread

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

80 points · 102 comments · theanonymousone

  1. Dinuda · · focus · HN ↗
    After 5.5, I basically don't notice a jump in model performance, other than my usage ending sooner.
    1. zero1009 · · focus · HN ↗
      Felt the same until I started using Luna. I feel like I get similar performance, but faster, and my usage lasts so much longer.
      1. Alifatisk · · focus · HN ↗
        Yup, GPT-5.6 Luna was a checkpoint for me. Too good for its price.
    2. ndbe · · focus · HN ↗

      [dead]

    3. tom1337 · · focus · HN ↗
      Kinda same but I miss my 5.3 Codex. Thing lasted forever on my $20 subscription and with detailed prompts was able to pretty much implement everything I requested it to do with a acceptable quality.
    4. ModernMech · · focus · HN ↗
      Same. 5.5 got work done then 5.6 was also fine then 6 was maybe not quite as good. Now with 6.1 they are cutting usage and raising prices and introducing ultra fast mode, but things were good enough 5 months ago.
    5. user43928 · · focus · HN ↗
      I do.

      The results are less buggy, animations are much better.

      It can work autonomously for hours and the result is decent most of the time.

      That wasn't usually the case with 5.5, which needed more feedback and iterations to get things right.

    6. baq · · focus · HN ↗
      Astra is noticeably smarter than any OpenAI model before it. Sol 6.1 is very noticeably smarter than sol 6 even after half a day of using it (sol 6 was actually terra 6 and opus 5.5 has taken them by a total complete surprise)
      1. rimliu · · focus · HN ↗
        none of them is smart.
        1. baq · · focus · HN ↗
          define smart
    7. whazor · · focus · HN ↗
      Meanwhile Fable 5 and Opus 5.5 are in a different league.
    8. redox99 · · focus · HN ↗
      Because you're probably using it for stuff like "edit this function"/"refactor this class". 5.5 to 5.6 Sol was a giant jump. 5.6 to 6.1 seems very large as well.
    9. kdaniel_03 · · focus · HN ↗
      It's the same for most tasks. Where I do notice it is long agent runs, where agents take more steps and the performance difference definitely compounds over the iterations.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.