‹ BackHN Continuity

Thread

LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents

403 points · 766 comments · Anon84

  1. stratos123 · · focus · HN ↗
    LeCun also said back in 2022 that "if you train a machine, as powerful as it could be, your 'GPT-5000', on text", it will never be able to learn basic common-sense physics like that objects placed on tables will move along with them.
    1. mdp2021 · · focus · HN ↗
      > never be able to learn basic common-sense physics

      And has it at this stage, within in-depth take of said "learning", foundationally?

      I have not been able to properly check the studies for a long time now, but I remain unaware of achieved solutions on the problem of reliably referencing a world model out of a language model - that "counting the 'r's in 'raspberry'" be not guessing, not memory, but actually counting.

      1. cavoirom · · focus · HN ↗
        counting 'r' in 'raspberry' to the LLM is similar to 4-dimension space to human. Their world's unit is token, not character, although they could use indirect method such as "run code" to find out. It will stay that way until they change the fundamental of the token that the LLM can perceive characters.
        1. estearum · · focus · HN ↗
          It's not even fair to call "run code" to be indirect compared to what a human would do. The word raspberry has no Rs in it in human language either. We have a written representation of it, which we can then write down either in our head or on paper, and then we can "run the algorithm" of counting each of the letters.

          Nothing intrinsically more or less direct about the LLM's method than ours.

          1. cavoirom · · focus · HN ↗
            I could argue LLM only have "token" as their perceivable dimension, compare to human multiple senses as the physic perceivable dimension and a brain with many other dimension of "learning" and "thinking". In spoken language, we may not have 'r' but in written we have, both spoken language and written language are learned skills.
            1. azornathogron · · focus · HN ↗
              Is "token" a directly perceivable unit for the LLM? If you ask it "how many tokens are in this sentence?" can it count them (again, not guessing or making a tool call)?

              I've never tried it and it might take some thought and effort to conduct an experiment to find out properly, but I would be interested in the answer.

              1. lawandjustice · · focus · HN ↗
                I dont think so. This is akin to asking a person, what is the frequency of the light hitting your eye when watching a leaf for example. You either know the (approximate) answer by knowing the frequency of green, or use a tool to measure it. If the LLM gives the correct answer it is either.guessing based on intution(and this intuition is based on pairs of word to tokenization length in text form in training data), writing code(or executing a tokenizer) or running a tokenizer mentally (reasoning via CoT).
              2. mdp2021 · · focus · HN ↗
                Not the point: the simulated intelligence in this context needs to create proper representation. It is not a matter of what it sees but of what it can see.
            2. scratcheee · · focus · HN ↗
              You could argue in return that humans only have electro-chemistry as our one perceivable dimension. We only indirectly perceive light through the signals our eyes send to our brains.

              In my mind agi is pretty much by definition a virtual machine, so the mechanisms behind thought are only relevant for the sake of efficiency (ie you can argue that LLMs make a poor basis for intelligence because tokens and natural language are a poor way to encode the world, but if you can run it on a big enough computer to counteract the inherent wasteful virtualisation then who really cares how it works under the hood?)

              1. cavoirom · · focus · HN ↗
                So LLM and human all have 1 dimenion perceivable signal, just LLM is 240p, and human is 8K in resolution, that's why we have 'r' in our signal, LLM still have 'r' in their signal, just because of the "low resolution", raspberry wasn't encoded with so many 'r' as in human signal.

                I will stop here before our analogies go too far.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.