‹ BackHN Continuity

Thread

GPT-6 Sol and Luna

1779 points · 855 comments · OfficialTurkey

  1. m_fayer · · focus · HN ↗
    I've been working with agents all year, but 5.6 Sol was some sort of sweet spot for me. Something about how it communicated verbally and its engineering instincts just clicked for me, and I was able to somehow predict it and jam with it. Like a colleague you click with. It's the first model I've gotten attached to. I'm concerned that whatever model supercedes it, while technically better, just won't feel quite as natural to work with. And this makes me feel very professionally vulnerable to the labs. I miss the days when my crucial tooling came from companies as reliable and predictable as, say, Jetbrains.
    1. redox99 · · focus · HN ↗
      Same. In fact I found 6 Astra to be a downgrade in situations where I didn't need the extra intelligence.
      1. cmrdporcupine · · focus · HN ↗
        Yeah.

        Astra was/is superior for planning type tasks. It was capable of doing seemingly magic things with rather vague/lazy instructions ("I need to be able to test this on Windows, maybe a qemu VM or something? Shrug." ... 1 hour later "yeah i built you a whole qemu + eval windows image + harness of powershell scripts + shell scripts to retrieve & verify harness.").

        And for UI work -- which is not something I do a lot of but do here and there -- it was clearly superior to 5.6 Sol.

        But it also feels sloppier? Somehow. And too expensive to use.

        We'll see how Sol 6 is.

        1. mavsman · · focus · HN ↗
          Glad you pointed out the UI work. I've been doing a lot of it and it's so much better than 5.6 as UI, it's unbelievable. I give it super ambiguous instructions and it's reading my mind. I do the same thing with 5.6 and I'm correcting it for a few minutes.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.