‹ BackHN Continuity

Thread

Towards Self-Driving Codebases

123 points · 100 comments · wilhelmklopp

  1. rrook · · focus · HN ↗
    My bet is that we'll see a second layer of harness emerge, as self-driving codebases become the target. There will be an application facing harness, orthogonal to the agent facing harness. The app harness will represent the software factory that is emergent for the specific application being developed.

    Anyway, here's mine, still wip:

    <a href="https:&#x2F;&#x2F;hale-lang.org&#x2F;docs&#x2F;dna&#x2F;" rel="nofollow">https:&#x2F;&#x2F;hale-lang.org&#x2F;docs&#x2F;dna&#x2F;

    <a href="https:&#x2F;&#x2F;github.com&#x2F;hale-lang&#x2F;hale&#x2F;issues&#x2F;690" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;hale-lang&#x2F;hale&#x2F;issues&#x2F;690

    1. cyanydeez · · focus · HN ↗
      I suspect it won&#x27;t be a harness, but just a more specific LLM trained in the universe of user-selected context of vetted resources.

      Why? Because LLMs are always going to be dumb when they&#x27;re trained at scale. Their ability to speak software diverges from their friendly user input layer. A harness won&#x27;t overcome that, but an LLM saddle ontop of a larger model would provide the type of feedback loops you&#x27;d want to look into.

      I don&#x27;t think you&#x27;ll find two deterministic systems will produce much.

      1. verdverm · · focus · HN ↗
        fine-tuning may be a more scalable approach to LLM personalization than sending all the same context to two LLMs

        I&#x27;m working towards both in my homelab to see which works better with little qwen

      2. rrook · · focus · HN ↗
        I think that dissolves the &quot;self driving&quot; distinction, though? The mechanisms for driving the codebase must be present in the codebase itself. Otherwise its just a regular out-of-band development process.
      3. cindyllm · · focus · HN ↗

        [dead]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.