‹ BackHN Continuity

Thread

Gemini 4 Argon

1699 points · 1187 comments · bradleyg223

  1. taylorfinley · · focus · HN ↗
    Ten days ago I had an experience with Gemini 3.8 flash that made me wonder if I was being routed to a different model under test. I was trying to use rocm with llama.cpp on my 128gb Strix Halo but could only get it to run Vulkan. I pasted the error message into agy and it proceeded to attach GDB to my GPU driver, reverse-engineer the kernel queue ioctl interface, and author an LD_PRELOAD C shim to get ROCm llama.cpp working on my Strix Halo. My jaw was hanging open the whole time.

    Edit to add the fix: <a href="https:&#x2F;&#x2F;gist.github.com&#x2F;birep&#x2F;6f2c8d490c7a29820997d57bd654c351" rel="nofollow">https:&#x2F;&#x2F;gist.github.com&#x2F;birep&#x2F;6f2c8d490c7a29820997d57bd654c3...

    1. amanguliani · · focus · HN ↗
      Can confirm - I am HEAVY claude user, but always like to check with AGY and CODEX in between. AGY with Gemini 3.8 flash cooked last couple of times and CODEX is basically out of the mix for me
      1. onlyrealcuzzo · · focus · HN ↗
        Sol 6.1 is quite good, but damn is it slow.

        I&#x27;m using it to run overnight tasks, and that&#x27;s it until my quota runs out.

        Canceled my subscription.

        1. unconscionable · · focus · HN ↗
          I find Opus 5.5 is better at giving a high level adversarial &quot;should you do this to begin with&quot; where GPT-6.1 is happy to go down any wrong path.

          Also canceled my ChatGPT Pro $200&#x2F;mo subscription. Their Oct 30 price hikes and slow GPT-6.1 model has me looking for alternatives.

          1. directdev · · focus · HN ↗

            [dead]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.