If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
It's _so_ good, I no longer bother with Sonnet and use it locally for everything.
Consider bumping reasoning down to Medium as a default though, I agree with simonw it over thinks <a href="https://simonwillison.net/2026/Aug/16/qwen-38-27b/" rel="nofollow">https://simonwillison.net/2026/Aug/16/qwen-38-27b/
Isn't there a way to set up a thinking budget so it automatically tells the model "Conclude your reasoning now and provide the final answer." ?
Xeoncross · · focus · HN ↗
Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.
It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https://github.com/kunchenguid/gnhf" rel="nofollow">https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.
jszymborski · · focus · HN ↗
Consider bumping reasoning down to Medium as a default though, I agree with simonw it over thinks <a href="https://simonwillison.net/2026/Aug/16/qwen-38-27b/" rel="nofollow">https://simonwillison.net/2026/Aug/16/qwen-38-27b/
felineflock · · focus · HN ↗