‹ BackHN Continuity

Thread

Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%

96 points · 42 comments · FKJ

  1. thallavajhula · · focus · HN ↗
    I've tried all of these and nothing really works. I have only 1 line in my CLAUDE.md file and that is "Always ground your responses." and that's it.

    Claude didn't care about it. When I pointed that out, it was apologetic and that was it.

    1. ChrisRR · · focus · HN ↗
      You may have better success by using a less ambiguous term. I've never heard of grounding in this context, so it may help to describe what you want more clearly

      Edit: I just asked Claude how it would interpret that and it said it could either mean that would not answer from memory alone and only anchor claims into things it can check, or it would tightly relate its responses to the context that I had supplied.

      If it chose the latter, I could see why it wouldn't always resort to search results

      1. WalterGR · · focus · HN ↗
        > I just asked Claude

        How well are LLMs able to reason about their own behavior?

        Put another way, this entire post is about getting AI to not bluff. How do you know that it’s not bluffing in its response to you?

        1. ChrisRR · · focus · HN ↗
          Well when you have a black box such as claude, the best you can do is ask it for its own interpretation. Anything else would just be a guess
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.