‹ BackHN Continuity

Thread

Don't be fooled–LLMs don't reason

75 points · 159 comments · leopoldj

  1. rkagerer · · focus · HN ↗
    I wonder if as a hack, some of the shortcomings mentioned could be addressed through prompting.

    E.g. "Approach this problem iteratively. As you form a hypothesis, track the confidence you have in various explanations you're considering, what evidence you're weighing to support each, and the unresolved questions you're holding onto. Log all that for later inspection.

    Be methodical when evaluating evidence and only accept facts you have verified. At every stage, gauge how much each possible next step resolves uncertainty, and discard options unlikely to advance progress. Divide the functions I described into subagents responsible for each, and coordinate with them as you work."

    1. vorticalbox · · focus · HN ↗
      You can sort of do this.

      For a given bug one could write a test that prove its existence, this gives the LLM a target that they can actually iterate towards.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.