GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Unofficial Hacker News client; not affiliated with Y Combinator.
minimaxir · · focus · HN ↗
This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.
joshstrange · · focus · HN ↗
Cache doesn't help you much when you are compacting every 5 minutes...
I was shocked at how quickly I ran out my $100/mo subscription with a single agent (sol medium).
onlyrealcuzzo · · focus · HN ↗
No LLM will be cost effective if it's compacting this often. You have to find a way around it.
ngruhn · · focus · HN ↗
sally_glance · · focus · HN ↗
SyneRyder · · focus · HN ↗
Apparently OpenAI makes you manually setup their 1 Million context window, and it seems to be only documented on X:
<a href="https://x.com/thsottiaux/status/2089082893804896524" rel="nofollow">https://x.com/thsottiaux/status/2089082893804896524
There's at least a forum thread about it here:
<a href="https://community.openai.com/t/why-does-codex-report-a-258-400-token-context-window-for-gpt-5-6-sol/1394346/5" rel="nofollow">https://community.openai.com/t/why-does-codex-report-a-258-4...
gf000 · · focus · HN ↗
rrvsh · · focus · HN ↗
kaoD · · focus · HN ↗
bjord · · focus · HN ↗
yes, exactly
SyneRyder · · focus · HN ↗
These are often my best sessions - they're unattended overnight, because by then we have the specification figured out, and I can just leave Claude to build out the rest, making good choices if it does find gaps in the spec. I regularly go to sleep & wake up to an entirely new application completed. Claude never uses compacting in my sessions.
I haven't used GPT as much as I should have, so I'm prepared to be incorrect & out of date. It just intuitively feels like I wouldn't get the same from a 275K context window - maybe it uses lots of subagents? Even Deepseek & GLM have 1 Million context windows now, so it "feels" strange for people to actually prefer the 275K window. But that's just my intuition.
bjord · · focus · HN ↗
if you talk about them (in which you lean on an LLM as a sort-of independent employee) and conservative, chunk-based usage (in which you use the LLM as more of an extension of yourself), you're comparing apples to oranges
a predefined spec obviously reduces that gap but how much is highly dependent on the level of detail
jeremyjh · · focus · HN ↗
I’ve also found compaction not to be a problem when it does happen.
threecheese · · focus · HN ↗
jeremyjh · · focus · HN ↗
onlyrealcuzzo · · focus · HN ↗
rrvsh · · focus · HN ↗
Benjamin_Dobell · · focus · HN ↗
~/.codex/config.toml