‹ BackHN Continuity

Thread

Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%

96 points · 42 comments · FKJ

  1. BatchJob · · focus · HN ↗
    The LLM will take a statistical path to reply and will not refuse to do so under any circumstances except where its been coded to do so.

    Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words.

    Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.

    1. astrange · · focus · HN ↗
      > LLMS dont understand the meaning of any words.

      In what way do you understand the meaning of the word "unicorn" that an LLM does not? It has experienced exactly as many real unicorns as you have.

      1. shiandow · · focus · HN ↗
        Experience is not understanding, but an LLM does not reason so it cannot understand.

        It can produce text that looks like reasoning. It can even produce text with mostly sound logic, but there is no internal experience or reasoning that occured there just the generation of language.

        LLMs therefore tend to be very bad at tasks that involve meta cognition. I've yet to successfully convince one to tell me when it knows something.

        1. serf · · focus · HN ↗
          I think the X cant do Y arguments require strict definitions of X and Y.

          w.r.t. this post : let's define 'reasoning' here, because there are definitions of 'reason' and 'reasoning' that would fit to a simple condition comparison let alone a massively complex llm.

          as for the meta cognition bit : show me a human that can accurately affirm when they know something. These kind of things aren't binary, nor can they be.

          1. otabdeveloper4 · · focus · HN ↗
            LLMs generate text, and they do not use any system of logic or syllogisms to do so. It's a pretty cut-and-dry obvsious statement of fact, no need to muddy the conversation here.
            1. allturtles · · focus · HN ↗
              So by your lights, the vast majority of humans don't reason either (and even those who do don't do it most of the time)?
              1. _g0xr · · focus · HN ↗
                Have you read the chinese room argument?

                <a href="https:&#x2F;&#x2F;rintintin.colorado.edu&#x2F;~vancecd&#x2F;phil201&#x2F;Searle.pdf" rel="nofollow">https:&#x2F;&#x2F;rintintin.colorado.edu&#x2F;~vancecd&#x2F;phil201&#x2F;Searle.pdf

                People have been having these discussions since 1980 at least. I&#x27;m not going to bother rehashing the issue endlessly in shortform internet comments, but this paper exists if you&#x27;re really interested in the subject.

                1. allturtles · · focus · HN ↗
                  Yep, familiar with it. I think the Chinese room is clearly a bad argument. The brain is just as much a &quot;chinese room&quot; as a computer is.
        2. astrange · · focus · HN ↗
          There is internal experience.

          <a href="https:&#x2F;&#x2F;www.anthropic.com&#x2F;research&#x2F;global-workspace" rel="nofollow">https:&#x2F;&#x2F;www.anthropic.com&#x2F;research&#x2F;global-workspace

          However, we don&#x27;t &#x2F;want&#x2F; them to have too much internal experience, because we want to know what they&#x27;re thinking* for safety reasons.

          * or, we want to be able to assume that the answer text is causally related to the thinking text

        3. Zambyte · · focus · HN ↗
          &gt; I&#x27;ve yet to successfully convince one to tell me when it knows something.

          This is a near daily experience for me when using a coding harness. I will ask it for some favts about the environment, and it will continuously explore the environment until it exhausts reasonable exploration, or it finds the facts.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.