‹ BackHN Continuity

Thread

Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%

96 points · 42 comments · FKJ

  1. thallavajhula · · focus · HN ↗
    I've tried all of these and nothing really works. I have only 1 line in my CLAUDE.md file and that is "Always ground your responses." and that's it.

    Claude didn't care about it. When I pointed that out, it was apologetic and that was it.

    1. ChrisRR · · focus · HN ↗
      You may have better success by using a less ambiguous term. I've never heard of grounding in this context, so it may help to describe what you want more clearly

      Edit: I just asked Claude how it would interpret that and it said it could either mean that would not answer from memory alone and only anchor claims into things it can check, or it would tightly relate its responses to the context that I had supplied.

      If it chose the latter, I could see why it wouldn't always resort to search results

      1. westurner · · focus · HN ↗
        > I've never heard of grounding in this context, so it may help to describe what you want more clearly

        An eval of this is likely worthwhile;

        Re: "Grounded in logic" and "Grounded in theory"

        Ground and justify all of the responses with logic and theory and real observations from qualified experiments with citations.

        Present a coherent argument borne of logical premises with extant sufficient proven evidence of support. Assess and critique the response given such criteria that all responses should be valid logical arguments, and revise before responding

        1. astrange · · focus · HN ↗
          The model, being smarter and more well-read than humans, is aware that what you ask is not possible.

          <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Logical_positivism#Decline_and_legacy" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Logical_positivism#Decline_and...

          Because LLMs also run off vibes and the writing style of your text, another important issue with your prompt here is that it makes you sound like a stuffy dork, or perhaps a pro se litigant. They won&#x27;t respond to this well because LLMs have feelings too.

          <a href="https:&#x2F;&#x2F;www.anthropic.com&#x2F;research&#x2F;emotion-concepts-function" rel="nofollow">https:&#x2F;&#x2F;www.anthropic.com&#x2F;research&#x2F;emotion-concepts-function

          Just be normal! And have evals.

          1. westurner · · focus · HN ↗
            So research methods like the scientific method are still subjective in AI implementation according to the AI experts?

            Once there are - or next month when there will be - better models, agents, and agent harnesses for this, do you think that then we should concisely specify what is required instead of doing evals for particular models?

            So meta-analysis and requisite language are too high-order for existing models and agents, and it&#x27;s currently necessary to apply such procedural controls outside of the prompt?

            1. astrange · · focus · HN ↗
              &gt; So research methods like the scientific method are still subjective in AI implementation according to the AI experts?

              I am not up to date on philosophy of science, but the scientific method is certainly always subjective, or at least can&#x27;t be successfully expressed in a formal system.

              Here&#x27;s a book you can read: <a href="https:&#x2F;&#x2F;metarationality.com" rel="nofollow">https:&#x2F;&#x2F;metarationality.com

              &gt; Once there are - or next month when there will be - better models, agents, and agent harnesses for this, do you think that then we should concisely specify what is required instead of doing evals for particular models?

              Hmm, not sure what you mean. &quot;Evals&quot; are another way of saying &quot;regression tests&quot;, so they&#x27;re useful when you want to change or compare any part of the system.

              &gt; and it&#x27;s currently necessary to apply such procedural controls outside of the prompt?

              In general I think you should try to move controls out of the prompt and into an external system, but the downside is that it costs more, so it&#x27;s not always necessary.

              1. westurner · · focus · HN ↗
                I&#x27;m aware of what evals are.

                Do you think it is wise to optimize prompts for specific models or agents when there is a new model every month?

                So, to build something like Co-Scientist the controls should be in the agent? Or RLHF&#x27;d like other things when training the model?

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.