If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
if you have 128 GB, you could use Qwen Flash Next at some reasonable quant, with the new SSD hack for only keeping some of the model resident in memory.
I have a Ryzen AI Max+ 395 with 128 GB for running those sort of "background task", and my sweet spot is currently Qwen3.8 Flash-Next IQ4 at 96 GB.
Xeoncross · · focus · HN ↗
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
ghilston · · focus · HN ↗
seanmcdirmid · · focus · HN ↗
nolok · · focus · HN ↗