Ten days ago I had an experience with Gemini 3.8 flash that made me wonder if I was being routed to a different model under test. I was trying to use rocm with llama.cpp on my 128gb Strix Halo but could only get it to run Vulkan. I pasted the error message into agy and it proceeded to attach GDB to my GPU driver, reverse-engineer the kernel queue ioctl interface, and author an LD_PRELOAD C shim to get ROCm llama.cpp working on my Strix Halo. My jaw was hanging open the whole time.
Edit to add the fix: <a href="https://gist.github.com/birep/6f2c8d490c7a29820997d57bd654c351" rel="nofollow">https://gist.github.com/birep/6f2c8d490c7a29820997d57bd654c3...
This is why I think llms are a killer app for Linux desktop. They’ve been trained on Linux very hard, and it cleanly solves the “how do I make it do $thing” problem since everything is open and the llm can manipulate it. For example: Sound not working? Just tell the llm.
Yes this might be a bit too noob linux but I have the AIs maintaining an env.md + logs / updates on that md whenever they change anything about the env. Even things as simple as getting tmux to be fast, they do an excellent job at.
Drivers, Coding environment setup etc. are great too and it's nice to have everything logged so the next (more powerful) agent can come and improve the thing once in a while.
taylorfinley · · focus · HN ↗
Edit to add the fix: <a href="https://gist.github.com/birep/6f2c8d490c7a29820997d57bd654c351" rel="nofollow">https://gist.github.com/birep/6f2c8d490c7a29820997d57bd654c3...
plasticchris · · focus · HN ↗
safog · · focus · HN ↗
Drivers, Coding environment setup etc. are great too and it's nice to have everything logged so the next (more powerful) agent can come and improve the thing once in a while.