For some reason they're comparing their harness running on Astra xhigh to Codex with Astra max?
---
Also worth noting that OpenAI just added support for async tool calling to their harness, which isn't 1:1 with this approach, but is slowly ramping up in being able to provide something similar.
A big part of why Codex uses so many tokens is that it basically hot loops on polling tasks it starts for... absolutely no good reason: <a href="https://www.reddit.com/r/codex/comments/1wdlp7q/weve_discovered_the_issue_behind_codex_harness/" rel="nofollow">https://www.reddit.com/r/codex/comments/1wdlp7q/weve_discove...
I fixed it on my fork of Codex too, also back in Jan/Feb – I keep this patch rebased, for anyone who wants it: <a href="https://github.com/tekacs/codex/commit/9ffcf8db9078eae43d4111ff94259795c1e962c9" rel="nofollow">https://github.com/tekacs/codex/commit/9ffcf8db9078eae43d411...
It results in token savings similar in scale to those displayed here by Unreal.
---
My harness has used a slightly fancier version of the approach that Unreal is using since ~Feb, and... it definitely works excellently, but it's also assuredly smoother with Astra and other recent models that are more aware of async tool calling.
Seems to do the job and reduce usage; I just ran Astra for ~5 hours (using a goal) and it used the last 30% of my usage. And now they released GPT-6 Sol and Luna (which is basically 5.6 Sol and Luna, but a bit better and also 50% cheaper) ;_;
Perfect, I will incorporate this as default as well as the commit from the other guy into my own codex fork <a href="https://github.com/AmazingTurtle/codex" rel="nofollow">https://github.com/AmazingTurtle/codex btw. I'm rebasing on 0.156.0 right now
tekacs · · focus · HN ↗
For some reason they're comparing their harness running on Astra xhigh to Codex with Astra max?
---
Also worth noting that OpenAI just added support for async tool calling to their harness, which isn't 1:1 with this approach, but is slowly ramping up in being able to provide something similar.
A big part of why Codex uses so many tokens is that it basically hot loops on polling tasks it starts for... absolutely no good reason: <a href="https://www.reddit.com/r/codex/comments/1wdlp7q/weve_discovered_the_issue_behind_codex_harness/" rel="nofollow">https://www.reddit.com/r/codex/comments/1wdlp7q/weve_discove...
I fixed it on my fork of Codex too, also back in Jan/Feb – I keep this patch rebased, for anyone who wants it: <a href="https://github.com/tekacs/codex/commit/9ffcf8db9078eae43d4111ff94259795c1e962c9" rel="nofollow">https://github.com/tekacs/codex/commit/9ffcf8db9078eae43d411...
It results in token savings similar in scale to those displayed here by Unreal.
---
My harness has used a slightly fancier version of the approach that Unreal is using since ~Feb, and... it definitely works excellently, but it's also assuredly smoother with Astra and other recent models that are more aware of async tool calling.
NostraDavid · · focus · HN ↗
AmazingTurtle · · focus · HN ↗