GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Unofficial Hacker News client; not affiliated with Y Combinator.
minimaxir · · focus · HN ↗
This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.
joshstrange · · focus · HN ↗
Cache doesn't help you much when you are compacting every 5 minutes...
I was shocked at how quickly I ran out my $100/mo subscription with a single agent (sol medium).
codewithcheese · · focus · HN ↗
redox99 · · focus · HN ↗
shimman · · focus · HN ↗
This is why these companies are struggling to make money, they're chastising their customers just like they've been chastising the human race.
trio8453 · · focus · HN ↗
It's very appropriate in the cases when you're holding it wrong. The fact that you're paying doesn't mean that you can't make mistakes or waste resources.
shimman · · focus · HN ↗
If this is how you want to get people on your side, I can understand why the entire country/human race are against these companies.
trio8453 · · focus · HN ↗
It's a product and if you're using it incorrectly, we can either
1. say so
2. pretend that you don't so to get/keep you on "our side"? or not say is because you're skeptical or hate it? (how does that last bit even follow logically?)
How is 2 better in any way for anyone involved?
crossroadsguy · · focus · HN ↗
Anonasty · · focus · HN ↗
jorblumesea · · focus · HN ↗
Aeolun · · focus · HN ↗
threecheese · · focus · HN ↗
I overused Astra in order to drain my weekly, figuring I'd have the reset. (not wastefully, I did get more work done)
Aeolun · · focus · HN ↗
seunosewa · · focus · HN ↗
nkmnz · · focus · HN ↗
edot · · focus · HN ↗
mattkenefick · · focus · HN ↗
I create a lot, but I can make a full month with Astra on the current Pro plan. What are you doing to spend that much?
redox99 · · focus · HN ↗
1 day is kind of generous, it probably lasts like 12 hours of running non stop.
apitman · · focus · HN ↗
antonvs · · focus · HN ↗
ChickeNES · · focus · HN ↗
Foobar8568 · · focus · HN ↗
Marha01 · · focus · HN ↗
antonvs · · focus · HN ↗
I've been using Gemini on development of a DNN training pipeline, and there's no way you can describe it as "dumb as hell". That description just makes it clear that you're not talking about the technical capabilities of the models, but about some sort of fanboy comparison from a parallel hype universe.
_davide_ · · focus · HN ↗
onlyrealcuzzo · · focus · HN ↗
No LLM will be cost effective if it's compacting this often. You have to find a way around it.
ngruhn · · focus · HN ↗
sally_glance · · focus · HN ↗
SyneRyder · · focus · HN ↗
Apparently OpenAI makes you manually setup their 1 Million context window, and it seems to be only documented on X:
<a href="https://x.com/thsottiaux/status/2089082893804896524" rel="nofollow">https://x.com/thsottiaux/status/2089082893804896524
There's at least a forum thread about it here:
<a href="https://community.openai.com/t/why-does-codex-report-a-258-400-token-context-window-for-gpt-5-6-sol/1394346/5" rel="nofollow">https://community.openai.com/t/why-does-codex-report-a-258-4...
gf000 · · focus · HN ↗
rrvsh · · focus · HN ↗
kaoD · · focus · HN ↗
bjord · · focus · HN ↗
yes, exactly
SyneRyder · · focus · HN ↗
These are often my best sessions - they're unattended overnight, because by then we have the specification figured out, and I can just leave Claude to build out the rest, making good choices if it does find gaps in the spec. I regularly go to sleep & wake up to an entirely new application completed. Claude never uses compacting in my sessions.
I haven't used GPT as much as I should have, so I'm prepared to be incorrect & out of date. It just intuitively feels like I wouldn't get the same from a 275K context window - maybe it uses lots of subagents? Even Deepseek & GLM have 1 Million context windows now, so it "feels" strange for people to actually prefer the 275K window. But that's just my intuition.
bjord · · focus · HN ↗
if you talk about them (in which you lean on an LLM as a sort-of independent employee) and conservative, chunk-based usage (in which you use the LLM as more of an extension of yourself), you're comparing apples to oranges
a predefined spec obviously reduces that gap but how much is highly dependent on the level of detail
jeremyjh · · focus · HN ↗
I’ve also found compaction not to be a problem when it does happen.
threecheese · · focus · HN ↗
jeremyjh · · focus · HN ↗
onlyrealcuzzo · · focus · HN ↗
rrvsh · · focus · HN ↗
Benjamin_Dobell · · focus · HN ↗
~/.codex/config.toml
AmazingTurtle · · focus · HN ↗
manmal · · focus · HN ↗
Gareth321 · · focus · HN ↗
It's crazy on Codex. I sometimes get just 2-3 turns before it compacts. It has forced me to use persistent project documentation for everything. Maybe that's not a bad thing but unless it reads all the documentation after every compaction (and uses half its cache), it goes off the rails. By comparison, Opus 5.5 is a breath of fresh air. It takes FAR longer to hit the cache limit and that means it keeps useful information in working memory far longer. I think this alone has resulted in a massive productivity and efficiency increase for me.
RugnirViking · · focus · HN ↗
exfalso · · focus · HN ↗
jmalicki · · focus · HN ↗
The longer your chat gets, the slower and more expensive it gets.
Subagents are expensive but they scale way closer to O(n) than O(n^2).
Have some agents make bug reports/feature requests/roadmaps (linear is very AI friendly), others coordinate, others work on grinding out an individual ticket.
If there is a good ticket-level description, it's a waste of time IMO to have a main agent do it, that should be an agent with fresh context that will do it better faster (the shorter the context, the better models are at using the context they're given).
jaktet · · focus · HN ↗
jmalicki · · focus · HN ↗
Whenever I see my main agent do a compaction, that to me is a clear sign I didn't have it delegate bounded tasks enough.
Still, I see no evidence Codex or Claude Code inherit full context of the main agent in subagents, I've always seen them be prompted, but this is something high priority on my list of unknowns to understand better...
KetoManx64 · · focus · HN ↗
Everyone else that uses memory files and a new conversation for each new sub project/feature rarely hit their weekly allotments.