Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
Unofficial Hacker News client; not affiliated with Y Combinator.
BatchJob · · focus · HN ↗
Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words.
Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.
SwtCyber · · focus · HN ↗
[dead]