Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Unofficial Hacker News client; not affiliated with Y Combinator.
BatchJob · · focus · HN ↗
Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words.
Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.
ben_w · · focus · HN ↗
AI are trained, not coded. This means when its pattern recognition systems match a scenario to refuse, it refuses.
Pattern recognition has always been a bit fuzzy.
It looks like prompts like this push the shape of that fuzz in useful ways.
> Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.
True.
Also applies to humans, but true nevertheless.
<a href="https://en.wikipedia.org/wiki/Münchhausen_trilemma" rel="nofollow">https://en.wikipedia.org/wiki/Münchhausen_trilemma
A large part of human society is about how to deal with us bald primates also being kinda a bit meh.
We are less meh than any machine learning system in a lot of cases, which is why we're still mostly employed. We're a bit more meh in a few narrower cases, however.