‹ BackHN Continuity

Thread

GPT-6 Sol and Luna

1779 points · 855 comments · OfficialTurkey

  1. jeffnash · · focus · HN ↗
    At this point, the deciding factors for me between Claude Code 20x and Codex Pro 20x are:

    1/ Usage limits: downstream of input/output cost, but resets and obscure windows and odd 20x plan / 5x plan != 4x usage math throw a wrench into it. Winner right now is Codex by a mile, especially when you factor in the fact that ChatGPT usage (even 6 Astra Pro) is essentially unmetered on the 20x plan. Always a bummer when asking if I should see a doctor about a rash means I can't code as much. It's also is a godsend if you use an MCP like oracle to automate the process of calling the Pro model on particularly tough problems, giving better planning results or deeper code analysis without burning usage.

    2/ Context window in the harness. Claude Code wins on this. There used to be a toml file workaround for Codex to extend the GPT context window to 1m, but this stopped working on the plans and only on per-token billing (ETA: noname120 pointed out this is no longer the case and it can be enabled again [1]). 252k is just not enough. Codex's compaction is very good, fwiw, but it happens so frequently that even a model as powerful as Astra sometimes loses the plot on long-running tasks.

    3/ Ability to use the plan outside of the official harness. Codex wins. Anthropic does shit like bills requests as extra usage if it sees a hermes.md in a commit.

    I've subscription hopped a bunch, and at times I've had both, but I keep coming back to Codex because it wins on 2/3.

    ETA: apparently I haven't been Keeping Up With the Altmans and new 20x signups have been disabled for a few weeks. I am grandfathered in, which makes the comparison above pretty much moot.

    [1]<a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49806060">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49806060

    1. paulmist · · focus · HN ↗
      &gt; Winner right now is Codex by a mile

      Opposite in my experience. I need to limit codex to 500k on medium&#x2F;low, still run out in 2-3 days with 1 CLI window. CC gives me 4-5 medium&#x2F;high days with 2-3 CLI windows, and Opus is still great for other regular dumb engineering&#x2F;refactoring.

      On the other hand my head starts to hurt if I read Opus for too long, hopefully they fixed it with 5.5.

      1. joshstrange · · focus · HN ↗
        This is my experience. After months of hearing how Codex limits were way higher I bumped to the $100&#x2F;mo plan after hitting my limits a day early on Claude due to some heavy usage + Fable (not normal for me, I often fit nicely in the $200&#x2F;mo plan). I hit the usage limit in a day with a single agent running on codex and the tiny context window was stifling. Yes, I&#x27;m comparing a $100 to a $200 plan but I extrapolated the usage (4x&#x27;d it) and it still wasn&#x27;t close, I got way more done with Opus.

        Using Agentsview (which might have it&#x27;s own issues) I was getting ~$200 of API usage in my 1 week Codex window (paid $100) vs ~$5,000 of API usage in 1 week for Claude (paid $200).

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.