‹ BackHN Continuity

Thread

HarnessTax: How Much Does the Harness Matter for Coding Agents?

233 points · 99 comments · matt_d

  1. spike021 · · focus · HN ↗
    Say I'm using Claude Code or GPT Codex's harnesses but also sending some queries to the respective Anthropic and OpenAI models via OpenRouter.

    Do harnesses and therefore sending the queries directly to the LLM providers have caching and other benefits that OpenRouter does not provide? Would I get any of those benefits if I simply proxied any requests to the major providers' harnesses through OpenRouter? Or only if the requests go straight from the harness to the provider's API?

    1. grave0x · · focus · HN ↗
      <a href="https:&#x2F;&#x2F;github.com&#x2F;grave0x&#x2F;statepod-open" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;grave0x&#x2F;statepod-open

      Its literally me realizing they don&#x27;t have a good solution for it. Context should be local. It saves money and potential LLM confusion.

      Plus i can foresee a bunch of other deterministic context and session handling being just plain unrealized due to the default we have currently

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.