If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
if you have 128 GB, you could use Qwen Flash Next at some reasonable quant, with the new SSD hack for only keeping some of the model resident in memory.
Xeoncross · · focus · HN ↗
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
ghilston · · focus · HN ↗
seanmcdirmid · · focus · HN ↗