Show HN: Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
Show HN: Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
PcChip · · focus · HN ↗
rancor · · focus · HN ↗
nullpoint420 · · focus · HN ↗
peddling-brink · · focus · HN ↗
I got excited about someone paying attention to intel. Oh well.
kamranjon · · focus · HN ↗
peddling-brink · · focus · HN ↗
gunalx · · focus · HN ↗
<a href="https://github.com/ggml-org/llama.cpp/blob/master/docs/backend/SYCL.md" rel="nofollow">https://github.com/ggml-org/llama.cpp/blob/master/docs/backe...
wronglebowski · · focus · HN ↗
peddling-brink · · focus · HN ↗
bitexploder · · focus · HN ↗
Also for the 3 people that ever read this and are curious about local models still, Qwen 27B 3.8 matched Sonnet 5 in the 17 DeepSWE tasks I have run so far, solving the exact same 7 it has. Caveat: datacurve combined low/medium/high Sonnet 5 data.
wingtw · · focus · HN ↗
bitexploder · · focus · HN ↗
dlcarrier · · focus · HN ↗
gunalx · · focus · HN ↗
In the end i took the sligth slowdown of vulkan to have a more stable and higher development velocity backend While being able to use identical setups on both intel and and gpus.
shayanjavadi · · focus · HN ↗
[dead]
aidiveyt · · focus · HN ↗
DylanMerigaud · · focus · HN ↗
antonyragleap · · focus · HN ↗
ContinuityLab · · focus · HN ↗
[dead]
kgeist · · focus · HN ↗
Usually people advertise "Go binary" when it's pure Go, not just a CGO wrapper. The project seems to be the result of 5 minutes of running Claude Code.
colinsane · · focus · HN ↗
so this is just a go binary that exec's `llama-server --model $MODEL_PATH ...`? there's room for llama wrappers, sure, but the readme only shows features that are already exposed directly from llama-server. it doesn't seem to do anything besides rename the CLI arguments: i don't get it.