Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Unofficial Hacker News client; not affiliated with Y Combinator.
thallavajhula · · focus · HN ↗
Claude didn't care about it. When I pointed that out, it was apologetic and that was it.
ChrisRR · · focus · HN ↗
Edit: I just asked Claude how it would interpret that and it said it could either mean that would not answer from memory alone and only anchor claims into things it can check, or it would tightly relate its responses to the context that I had supplied.
If it chose the latter, I could see why it wouldn't always resort to search results
WalterGR · · focus · HN ↗
How well are LLMs able to reason about their own behavior?
Put another way, this entire post is about getting AI to not bluff. How do you know that it’s not bluffing in its response to you?
ChrisRR · · focus · HN ↗