‹ BackHN Continuity

Thread

Pentagon says overreliance on AI contributed to missile strike on Iran school

974 points · 553 comments · devonnull

  1. legitster · · focus · HN ↗
    > It found the U.S. “failed in its obligation to do everything feasible to verify” that the school was a military objective and that the failure “went beyond mere negligence.” The report said the United States “directed the strikes at the building of the school while being aware of a substantial risk of striking a civilian object and acting recklessly as regards the possibility that this would happen.”

    Reading the details, "AI" doesn't really seem like the culprit -it's a scapegoat.

    The intelligence that it was no longer a military target never entered the target database, the team that was responsible for vetting the target list was gutted, and said team was never even consulted.

    The White House wanted 1000 targets and pulled from their database without any due diligence. Whether it was an AI call or an SQL query - this was from pure human maliciousness and incompetence.

    1. xlayn · · focus · HN ↗
      Now, ask yourself if you want AI taking decisions around your life... it will pull numbers... someone will have a "goal" of 1000 "hits" and there you go, 2-3 months in jail working to check what went wrong...

      But don't worry, the other P AI system will ensure that the first one doesn't do mistakes...

      1. toomuchtodo · · focus · HN ↗
        Lawsuit Says UnitedHealth Used AI With a 90% Error Rate to Deny Care - <a href="https:&#x2F;&#x2F;www.yahoo.com&#x2F;news&#x2F;us&#x2F;articles&#x2F;lawsuit-says-unitedhealth-used-ai-173017952.html" rel="nofollow">https:&#x2F;&#x2F;www.yahoo.com&#x2F;news&#x2F;us&#x2F;articles&#x2F;lawsuit-says-unitedhe... - September 18th, 2026

        UnitedHealth uses AI model with 90% error rate to deny care, lawsuit alleges - <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=38299921">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=38299921 - November 2023 (58 comments)

        <a href="https:&#x2F;&#x2F;litigationtracker.law.georgetown.edu&#x2F;wp-content&#x2F;uploads&#x2F;2023&#x2F;11&#x2F;Estate-of-Gene-B.-Lokken-et-al_20231114_COMPLAINT.pdf" rel="nofollow">https:&#x2F;&#x2F;litigationtracker.law.georgetown.edu&#x2F;wp-content&#x2F;uplo...

        1. teravor · · focus · HN ↗
          that smells like bullshit to be fair. if you have a system that makes errors over 90% of the time then you have a system you can leverage to make fewer errors.

          the situation would have to be very particular&#x2F;contrived for that not to be the case.

          1. toomuchtodo · · focus · HN ↗
            <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;The_purpose_of_a_system_is_what_it_does" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;The_purpose_of_a_system_is_wha...
            1. 1718627440 · · focus · HN ↗
              The word purpose exists to describe something that is independent of it&#x27;s current behaviour. If you do not mean that, don&#x27;t use that word.
          2. flossly · · focus · HN ↗
            What if you want errors to happen, e.g. deny care, bomb civilians; in that case 90% error rate is what you want, and plausible deniably so!
          3. ImPostingOnHN · · focus · HN ↗
            What makes you think they are financially motivated to make fewer &quot;errors&quot; (improper denial of care while keeping premiums)?
          4. degamad · · focus · HN ↗
            &gt; the situation would have to be very particular

            The situation is very particular. When the system makes an error, I make a bigger profit. If the system made fewer errors, it would approve more things that cost me money.

            Why do you think I&#x27;m motivated to make fewer errors under that situation?

          5. dylan604 · · focus · HN ↗
            It sounds like something with a 90% error rate would be better to just do the opposite of what was determined and suddenly have a 10% error rate??
            1. eks391 · · focus · HN ↗
              That&#x27;s not how type 1 &amp; type 2 errors work. The article is describing type 2 (false negative) and inverting the output has no correlation with the percent of false positives (type 1)

              <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Type_I_and_type_II_errors" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Type_I_and_type_II_errors

          6. benlivengood · · focus · HN ↗
            Nah, it&#x27;s cumulative errors per claim. E.g. maybe the per-decision-factor rate is somewhere over 50% but overall most claims will have errors because so many factors go into coverage&#x2F;reimbursement decisions.
          7. Dylan16807 · · focus · HN ↗
            &gt; if you have a system that makes errors over 90% of the time then you have a system you can leverage to make fewer errors.

            Sure if you&#x27;re flipping a coin.

            If you&#x27;re calculating a number from 1 to 100 then depending on distribution a calculation that&#x27;s wrong 90% of the time is probably useless.

          8. mschuster91 · · focus · HN ↗
            &gt; if you have a system that makes errors over 90% of the time then you have a system you can leverage to make fewer errors.

            member the three words &quot;delay, deny, depose&quot;? Insurances - and not just in the US, Germany has similarly bad stuff going on - aren&#x27;t making a profit if they just blindly accept claims. It is much more profitable to deny with a vague &quot;AI&quot; system (or a blanket automated deny) first, and only have a first look with an actual human agent at it when the customer complains or files a lawsuit.

          9. j4yav · · focus · HN ↗
            I can easily build you a system that will guess the price of a stock tomorrow wrong at least 90% of the time. Even 100% wrong if you&#x27;d like. How would you use that to make fewer errors in some other system?
            1. [deleted] · · focus · HN ↗

              [deleted]

            2. 1718627440 · · focus · HN ↗
              If it is a binary choice you can just do the opposite.
              1. j4yav · · focus · HN ↗
                What’s binary choice we are discussing here and what’s the opposite that you’d do? Approve every claim that was denied by the AI, and vice versa? Or what’s the binary choice related to stock prices you’re referring to?
                1. 1718627440 · · focus · HN ↗
                  I agree, that most real world problems, are not binary choices. But in this case actually yes, just do the opposite.

                  For that reason error rates above 50% don&#x27;t make any sense. Building a system, that is reliably wrong, is as hard as building one that is reliable correct, because the systems are the same, just with the output inverted.

                  The worst you can get with a system is 50%, at which point you can just flip a coin, because that is just pure random. Any deviation from that is going to be an improvement.

                  1. [deleted] · · focus · HN ↗

                    [deleted]

                  2. j4yav · · focus · HN ↗
                    I always thought insurance claims were more complex, some things may be wholly or partially covered on a line by line basis, and there should be justification provided. In certain cases they go to mediation or even court. It doesn’t seem binary to me but maybe it works differently where you are.
                    1. 1718627440 · · focus · HN ↗
                      I said &quot;if it is&quot;, not that it would be.
                      1. j4yav · · focus · HN ↗
                        Oh, haha, sure.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.