‹ BackHN Continuity

Thread

When did Google get so weird?

2011 points · 1123 comments · sancho-panza

  1. Hugsbox · · focus · HN ↗
    Yesterday I tried to google "can the Halifax Wanderers still make the CPL playoffs?"

    So obviously what appears right at the top is the AI summary, which told me "they've already secured their #4 position and made the playoffs". I knew this wasn't true, and I guess I could have just scrolled down a bit further and found my answer but now I was curious.

    So I said "that's not true, they're still #5, what I want to know is _could they still make the playoffs_"

    It says they've got an upcoming game against Ottawa, and if they win their chances are good. That game has already taken place, so I correct it again and finally I get a reasonable answer.

    My question is: what's the point of the AI in the search engine if it itself isn't going to use the search engine first before answering? Like, I can't wrap my head around that. The answer is on the same page as its hallucination. It could have done a cursory look around before first hallucinating something completely false, and when corrected the first time giving me outdated information. It's meant to be A SEARCH ENGINE!

    1. beloch · · focus · HN ↗
      This is similar to how, not too long ago, LLM's had extreme difficulty counting the number of letters in some words. LLM's don't "think" or "reason" in the normal definition of those terms. They can do some pretty amazing things, but still screw up basic things like telling you something that is obviously wrong and contradicts the top search results.

      LLM's, in their present stage of development, are sort of like a crack-addled idiot savant. Sometimes they are obviously insane, and sometimes they seem quite cogent, but you must never trust them implicitly. This may be why they are so difficult to constrain. You could give them something equivalent to the laws of robotics, but following laws requires thought processes they simply don't have.

      I'm actually sort of amazed Google doesn't make people accept some kind of butt-covering EULA and post disclaimers about the inaccuracy of results before even showing you their AI's output. Are they not being sued over this kind of thing?

      1. VCFundedGenYer · · focus · HN ↗
        LLMs still can't do math nor count letters in words. Nothing has changed there.
        1. walrus01 · · focus · HN ↗
          This is true but a sufficiently smart LLM (run in a harness like opencode, no special MCP, no customization done whatsoever) will quickly turn out a basic 1 to 2 page sized python script to do the math. They can't do the math with any guarantee of accuracy with their own internal reasoning since it's a language model.

          But, for example, if you ask deepseek v4 flash 0731 to produce a python script to calculate the distance or azimuth directions between two points on an oblate spheroid using the vincenty and haversine geodetic formulas, it'll turn out the factually accurate vincenty and haversine formulas which has a perfect 100% correlation with what is hard coded into human-written GIS software. These things are clearly in its training data set from whatever whole-internet-crawl/scrape built the training set.

          Heck, just for fun I asked a reasonably smart LLM to re-implement the Karney formula (which is considerably more complex than Vincenty), just in case I ever had a need to calculate the distance between two points down to the nanometer, and it did it: <a href="https:&#x2F;&#x2F;www.google.com&#x2F;search?&amp;q=karney+formula+geodetic+" rel="nofollow">https:&#x2F;&#x2F;www.google.com&#x2F;search?&amp;q=karney+formula+geodetic+

          reference: <a href="https:&#x2F;&#x2F;github.com&#x2F;pbrod&#x2F;karney" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;pbrod&#x2F;karney

          You still have to be skeptical of its results and capable of understanding if it&#x27;s gone off on a hallucinatory path, but saying LLMs can&#x27;t do math isn&#x27;t really a hundred percent accurate anymore. More precisely it&#x27;s that they can&#x27;t do the math internally but they&#x27;re quite capable of producing the tool that does the math. And often producing a basic one-off tool that does the math takes less than a few seconds, then it runs it, and will spit back the results.

          Deepseek v4 flash 0731 (a somewhat randomly chosen example) isn&#x27;t even particularly sophisticated, large, or capable compared to a GLM5.3 size model or Kimi K3 size thing.

          1. AdieuToLogic · · focus · HN ↗
            &gt; Heck, just for fun I asked a reasonably smart LLM to ...

            LLMs are neither smart nor stupid. They are statistical token generators whose results are dependent upon their training data set and involve a degree of randomness.

            &gt; You still have to be skeptical of its results and capable of understanding if it&#x27;s gone off on a hallucinatory path ...

            Again, LLMs do not &quot;hallucinate.&quot; They are statistical token generators whose results are dependent upon their training data set and involve a degree of randomness.

            Nothing more.

            See also anthropomorphism[0].

            &gt; More precisely it&#x27;s that [LLMs] can&#x27;t do the math internally but they&#x27;re quite capable of producing the tool that does the math.

            This still falls under the purvey of statistical token generation. To wit, given enough variations of:

              bc -e &#x27;1 + 2&#x27;
              bc -e &#x27;41 + 1&#x27;
              ...
            
            LLMs can identify the addition expression in &quot;What is 4 + 1?&quot; and then emit a `&#x27;bc &quot;4 + 1&quot;&#x27;` command to produce a response. This is not &quot;doing&quot; or &quot;understanding&quot; math.

            It is pattern recognition, a task in which ANNs[1] excel.

            0 - <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Anthropomorphism" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Anthropomorphism

            1 - <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Neural_network_(machine_learning)" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Neural_network_(machine_learni...

            1. hodgehog11 · · focus · HN ↗
              During conversation, we are statistical token generators whose results are dependent upon our training set. Seriously, write that definition out rigorously. It encompasses virtually everything. It is totally meaningless. So to say &quot;nothing more&quot; is effectively also a tautology.

              This argument was asinine in 2024. It is insane to be saying these things in 2026. Where have you been? What have you been looking at? How many articles explaining why the &quot;statistical parrot&quot; analogy fails have you missed? How much mental gymnastics do you have to do to explain how a modern LLM can solve novel math problems that fall really far outside of its training set?

              It absolutely understands how to do math, by whatever reasonable definition you want to provide to the word &quot;understand&quot;. For example, the identification of the addition expression is understanding, and no, it does not do tool calling for basic arithmetic any more than humans might. Isolation of individual concepts in intermediate layers can already be demonstrated, or else transfer learning wouldn&#x27;t possibly work. Nobody is saying that LLMs are humans. But we need labels for some of the things that we observe and dismissing them because &quot;statistical&quot; is laughable.

              Look at the proof of this: <a href="https:&#x2F;&#x2F;github.com&#x2F;anthropics&#x2F;formal-math&#x2F;blob&#x2F;795efb86f191735c5481675763537cfb4ff37e55&#x2F;percolation&#x2F;summary.pdf" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;anthropics&#x2F;formal-math&#x2F;blob&#x2F;795efb86f1917... . Forget the Lean, look at the underlying argument construction. At the very least, this is continuing from an argument that was hinted at in the literature in 2024, but these proceedings were difficult enough that humans were not able to do them within two years. Do you attribute this to the harness alone? If so, that&#x27;s a pretty sophisticated bit of engineering, I would say! Probabilities are far too small to argue infinite monkey theorem.

              If there was even a shred of a reasonable argument that LLMs were incapable of concept extraction and manipulation, I and my colleagues would be all over it. We would relish in it. It would bring us comfort. It is unbelievable that people think they can spew whatever basic garbage they think of as a gotcha, and think that minds all over the world haven&#x27;t already considered that. This is like climate denial at this point.

              1. AdieuToLogic · · focus · HN ↗
                &gt; During conversation, we are statistical token generators whose results are dependent upon our training set. Seriously, write that definition out rigorously.

                If you do not see a difference between humans conversing (known consciousness as defined by humans) and the output of an LLM (known algorithms as defined by humans), I don&#x27;t know what to say.

                1. hardbass · · focus · HN ↗
                  Do you believe in souls?
                  1. latentsea · · focus · HN ↗
                    Have you ever seen one?
                    1. hardbass · · focus · HN ↗
                      Haven&#x27;t yet seen evidence for any.
                    2. butlike · · focus · HN ↗
                      Yes, the numerical point counter at the bottom of the popular video game Dark Souls. I doubt it was the soul anyone was expecting, but they do, in fact, exist.
                      1. diseasedyak · · focus · HN ↗
                        I like this reply.
                      2. latentsea · · focus · HN ↗
                        To be fair my first was when I played Soul Reaver. The most recent was in Minecraft Dungeons.
                2. epihelix · · focus · HN ↗
                  &gt; If you do not see a difference between humans conversing (known consciousness as defined by humans) and the output of an LLM (known algorithms as defined by humans), I don&#x27;t know what to say.

                  <a href="https:&#x2F;&#x2F;www.pnas.org&#x2F;doi&#x2F;abs&#x2F;10.1073&#x2F;pnas.2524472123" rel="nofollow">https:&#x2F;&#x2F;www.pnas.org&#x2F;doi&#x2F;abs&#x2F;10.1073&#x2F;pnas.2524472123

                  Whatever you might think about your own abilities, most individuals can&#x27;t tell the difference.

                  1. AdieuToLogic · · focus · HN ↗
                    &gt;&gt; If you do not see a difference between humans conversing (known consciousness as defined by humans) and the output of an LLM (known algorithms as defined by humans), I don&#x27;t know what to say.

                    &gt; Whatever you might think about your own abilities, most individuals can&#x27;t tell the difference.

                    I have yet to see an LLM say &quot;hello&quot; to a neighbor. I have done so and can definitively assure you &quot;most individuals&quot; can tell the difference.

                    1. thereforegrin · · focus · HN ↗
                      do you have a neighbour LLM who does not say &quot;hello&quot; when it sees you and THAT is how you know it&#x27;s an LLM?

                      I&#x27;m a bit confused by your argument because I too have some neighbors who don&#x27;t say &quot;hello&quot; when they see me. Are they LLMs too, you think?

                      Consciousness is a thing we assume of others because of tact not fact.

                3. redsocksfan45 · · focus · HN ↗

                  [dead]

                4. hodgehog11 · · focus · HN ↗
                  You are responding to a claim about mathematical definitions with subjective experience. No, consciousness is not well-defined.
                5. spider-mario · · focus · HN ↗
                  That’s not what they said. They said that the difference is not that.
                  1. AdieuToLogic · · focus · HN ↗
                    For context, in response to my original statement:

                      [LLMs] are statistical token generators whose results are 
                      dependent upon their training data set and involve a degree 
                      of randomness.
                    
                    This is literally what was written:

                      During conversation, we are statistical token generators 
                      whose results are dependent upon our training set.
                    
                    &gt;&gt; If you do not see a difference between humans conversing (known consciousness as defined by humans) and the output of an LLM (known algorithms as defined by humans), I don&#x27;t know what to say.

                    &gt; That’s not what they said.

                    How did I misquote and&#x2F;or mischaracterize any the above?

                    1. spider-mario · · focus · HN ↗
                      In that pointing out that “A can’t be X, unlike B, because A is Y” is fallacious if B is also Y does not entail that A and B can’t be different in other respects?

                      Hypothetical you: “Bread is neither tasty nor disgusting (unlike maple syrup). It’s a bunch of molecules.”

                      Hypothetical hodgehog: “Maple syrup is also a bunch of molecules [so if you accept that maple syrup can be delicious, being a bunch of molecules can’t be why bread isn’t].”

                      Hypothetical you: “If you don’t see a difference between bread and maple syrup, I don’t know what to say.”

                      1. hardbass · · focus · HN ↗
                        He quoted something from the Bible, so its possible he is a Christian and well theology could cloud clear thinking on matters of consciousness due to the soul stuff.
                        1. AdieuToLogic · · focus · HN ↗
                          &gt; He quoted something from the Bible ...

                          It is also possible that the proverb I provided to you is the origin of the oft quoted:

                            Better to remain silent and be thought a fool than to speak 
                            and remove all doubt.
                          
                          So there&#x27;s that.
                          1. hardbass · · focus · HN ↗
                            That applies very well to you. I asked you a very simple question and you failed to answer and are instead posting random quotes.
                6. hardbass · · focus · HN ↗
                  Then disprove the physical Church Turing hypothesis in regard to the human brain.
                  1. AdieuToLogic · · focus · HN ↗
                    &gt; Then disprove the physical Church Turing hypothesis in regard to the human brain.

                    The onus is not mine to disprove a hypothesis you have chosen to mention in passing. The responsibility is yours to prove said hypothesis or at least contribute meaningfully with some amount of credible research.

                    Or try to learn from Proverbs 17:28[0]:

                      Even fools are thought wise if they keep silent, and 
                      discerning if they hold their tongues.
                    
                    Either works for me.

                    0 - <a href="https:&#x2F;&#x2F;www.biblegateway.com&#x2F;passage&#x2F;?search=proverbs%2017:28&amp;version=NIV" rel="nofollow">https:&#x2F;&#x2F;www.biblegateway.com&#x2F;passage&#x2F;?search=proverbs%2017:2...

                    1. hardbass · · focus · HN ↗
                      The onus is on you to disprove a hypothesis that has as of yet held up to all of known physics, not put a random bible quote that has no relation to the conversation and a complete failure to answer a basic question.
              2. butlike · · focus · HN ↗
                Does the Robin bird understand the worm it&#x27;s pecking at? Honestly I find your comment asinine and overly aggressive.
                1. hodgehog11 · · focus · HN ↗
                  Yes, of course it was aggressive. It is frustrating to experience so many armchair experts on a forum usually populated with intelligent people, regurgitating debunked arguments from years ago, which get in the way of educating people about what is really going on. See the recent Hoog video for how frustrating this is. I believe this is how the climate scientists felt.

                  And yes, according to our best definitions, the Robin bird does understand the worm it&#x27;s pecking at.

              3. wartywhoa23 · · focus · HN ↗
                &gt; It absolutely understands how to do math

                Bout of tinnitus, then crickets

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.