‹ BackHN Continuity

Thread

Gemini 4 Argon

1699 points · 1187 comments · bradleyg223

  1. taylorfinley · · focus · HN ↗
    Ten days ago I had an experience with Gemini 3.8 flash that made me wonder if I was being routed to a different model under test. I was trying to use rocm with llama.cpp on my 128gb Strix Halo but could only get it to run Vulkan. I pasted the error message into agy and it proceeded to attach GDB to my GPU driver, reverse-engineer the kernel queue ioctl interface, and author an LD_PRELOAD C shim to get ROCm llama.cpp working on my Strix Halo. My jaw was hanging open the whole time.

    Edit to add the fix: <a href="https:&#x2F;&#x2F;gist.github.com&#x2F;birep&#x2F;6f2c8d490c7a29820997d57bd654c351" rel="nofollow">https:&#x2F;&#x2F;gist.github.com&#x2F;birep&#x2F;6f2c8d490c7a29820997d57bd654c3...

    1. danpalmer · · focus · HN ↗
      3.8 Flash is my daily driver and produces pretty excellent results all round.
      1. dcl · · focus · HN ↗
        What harness you use? Have you tested more than 1?
        1. danpalmer · · focus · HN ↗
          This experience is with Antigravity both internally and externally, and I have done quite a few side-by-side comparisons with the same prompt across a number of different Google and non-Google models.

          I&#x27;ve tried Codex as a harness too, and that was nice. I don&#x27;t find a significant difference between Antigravity and Codex. Codex has more features but I don&#x27;t use them.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.