‹ BackHN Continuity

Thread

US Military had close call after using AI for hallucinated intelligence report

519 points · 396 comments · realsarm

  1. drtgh · · focus · HN ↗
    > relatively poorly understood technology

    Poorly understood? how convenient...

    LLMs are vectorial databases with losses that index statistically filled data, which uses a text interface to query such statistically filled data. The output is a string concatenation (statistically concatenated bit by bit).

    When the LLMs are queried (prompted), you can get random mixed data as output, ERRORS, due to undesired indexes getting closer at one point while the string was being concatenated for the output, what affects the rest of the indexed content that will be concatenated.

    It is intrinsic to this tech. The larger the context, the greater the probability of get mixed data. And if the provider lowers the precision of those indexes -in order to decrease hardware resources and energy consumption- such probability increases to the point where those errors are granted.

    Even knowing that the queries can return wrong/mixed data in the responses, errors, the companies developing this, decided to introduce a new product, that connects such LLMs outputs to the command console, latter connected to internet, raw 'eval' running commands from such outputs witch obviously can contain whatever mixed random. Then we started to hear "oh, it deleted my directory", etc, and it seems the next one will be "a missile killed my wife", because it is a text concatenation engine with errors.

    To name it "hallucination" is an euphemism... those are errors, and they are granted to happen at one moment. If they do not know this, then they ate too much marketing without doing their job, or it was a convenient contract for the pocket$ of someone.

    1. margalabargala · · focus · HN ↗
      I agree with most of your comment, but...

      > To name it "hallucination" is an euphemism... those are errors

      I find this and other "don't anthropomorphize the computer" statements incredibly unconvincing.

      People develop terms for things and language has always contained overloaded or "literally inaccurate" terms.

      An LLM can have "hallucinations" in the same way a modern computer program can have "bugs".

      1. usernomdeguerre · · focus · HN ↗
        I disagree, I think 'Hallucination' is a risk-shedding weasel-word. It's meant to shift blame away from the technology and its creator (multibillion dollar AI companies etc) in a way that doesn't hold those actors accountable or responsible for the outcomes.

        In any other software it would be an error, regression, bug. And in a human process it would be at ~least something someone would call 'bullshit'.

        1. vorticalbox · · focus · HN ↗
          I’m not sure either would is particularly good at describing what is happening.

          Error in implies something broke, which nothing broke the LLM did exactly what they where designed to do generate text based on a statistically likely bases.

          Hallucination Does really fit here either. It implies it’s experiencing something that is not there which it isn’t experiencing anything.

          1. Towaway69 · · focus · HN ↗
            Howabout: lied. The LLM lied indirectly (perhaps) but it made a claim that was false. Which is a lie.

            Humans lie and LLMs “hallucinate”? What gives. It’s an untruth that the LLM is selling for a truth, that’s lying in my books.

            And since we don’t know how or why the LLM works, we can’t even judge whether it explicitly lied or only because it didn’t know better.

            1. vorticalbox · · focus · HN ↗
              Lie implies it knows what it is saying to be false.
          2. t-3 · · focus · HN ↗
            Unexpected Result is perhaps a more accurate description.
        2. segsegsgsg · · focus · HN ↗
          error, regression, bug, bullshit are not weasel words, hallucination is a weasel word because why exactly? your argument is a weasel argument.
          1. usernomdeguerre · · focus · HN ↗
            I explained why, if you want me to engage i'll try but you're asking me to restate my position.
            1. s1artibartfast · · focus · HN ↗
              You made a lot of clams about how it intentionally shifts blame, but none of them are supported.

              Do you really think accountability would be meaningfully different had they been called bugs?

              Why are you so confident it is intentional? My understanding of the history is that it was a technical term among researchers long before it had any public mind share. It's popular because it's and intuitive for most people, not because there was a concerted effort hooked up by some PR and legal team.

            2. e5yyey · · focus · HN ↗
              I disagree, I think 'Bug' is a risk-shedding weasel-word. It's meant to shift blame away from the technology and its creator (multibillion dollar AI companies etc) in a way that doesn't hold those actors accountable or responsible for the outcomes. In any other software it would be an error, regression, hallucination. And in a human process it would be at ~least something someone would call 'bullshit'.
        3. leonidasrup · · focus · HN ↗
          Using the term "Hallucination" makes it sound less problematic, less impactfull for user.

          They should have used the term "error". For example in statistics, there many kinds of errors, discretization error, prediction error, sampling error, ...

          <a href="https:&#x2F;&#x2F;www.statisticshowto.com&#x2F;errors-in-statistics&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.statisticshowto.com&#x2F;errors-in-statistics&#x2F;

          1. lukan · · focus · HN ↗
            I don&#x27;t know, but all humans make errors, but if some humans are known to hallucinate, you don&#x27;t let them do important things unsupervised. So error sounds actually less problematic to me.
            1. verdverm · · focus · HN ↗
              We do a surprising amount of &quot;hallucinations&quot; without the extreme version of hallucinations. We assemble things we sort of remember into incorrect statements all the time. I&#x27;m sure every one of us has been corrected for misremembering something or stating something based on misremembered facts (plague of clickbait headlines).

              This is more or less how I see the LLM output, but as a path finding exercise over next-token probability graphs. This is (i.e.) why they are trained to use phrases like &quot;wait but&quot; or &quot;actually&quot;, these words even out the probability of different paths, giving them their ability to &quot;consider&quot; different solutions.

            2. leonidasrup · · focus · HN ↗
              Because so few humans hallucinate, many imagine hallucinations like dreams and dreams are mostly harmless.
        4. plant-ian · · focus · HN ↗
          Totally weasel words in this time frame. I think in 10 years after everyone has a better understanding of what we are dealing with these weasel words would maybe make sense. Right now it seems more sensible to deem this at best a false positive, or glitch, or if it must be anthropomorphized a screw up or a f&#x27; up. I don&#x27;t think the llms are dehydrated. Although that&#x27;s funny on another level. Edited: to be less abrasive
        5. margalabargala · · focus · HN ↗
          &gt; In any other software it would be an error, regression, bug.

          How is &quot;bug&quot;, literally an organism with a will of its own that you cannot control, any less of a weasel word?

          1. usernomdeguerre · · focus · HN ↗
            <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Bug_(engineering)#History" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Bug_(engineering)#History

            But on reflection I don&#x27;t disagree it was probably made for similar effect in the era of human software development. That sounds like it strengthens my point?

            1. margalabargala · · focus · HN ↗
              On the contrary I think it weakens your point.

              People don&#x27;t consider &quot;bug&quot; a weasel word, to the point that you yourself held it up as an example of not being a weasel word, despite it being a willful, uncontrollable organism.

              I see no reason why &quot;hallucination&quot; won&#x27;t become a similar piece of neutral jargon. It already is for many people, even if you&#x27;re not (yet?) among them.

              1. usernomdeguerre · · focus · HN ↗
                It becomes a neutral term because we are practicioners (presumably?) who benefit from it and have thus let it become habitual. I already admitted it was weasely upon reflection.

                If you want this class of LLM error to also become habitual and neutral then fine, I don&#x27;t, and I think many others don&#x27;t.

          2. elzbardico · · focus · HN ↗
            You need to consult the etymology of the word bug in the context of computing. In the pre-historical times of vacuum-tube computers, with exposed conductors with potentials sometimes over 200v, a coachroach in the wrong place can definitely flip a few bits in the most benign scenario.
            1. margalabargala · · focus · HN ↗
              Right, yes, that&#x27;s my point...you need to consult the comment you replied to.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.