LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
Unofficial Hacker News client; not affiliated with Y Combinator.
stratos123 · · focus · HN ↗
mdp2021 · · focus · HN ↗
And has it at this stage, within in-depth take of said "learning", foundationally?
I have not been able to properly check the studies for a long time now, but I remain unaware of achieved solutions on the problem of reliably referencing a world model out of a language model - that "counting the 'r's in 'raspberry'" be not guessing, not memory, but actually counting.
ethbr1 · · focus · HN ↗
Incredibly inefficiently because of the recursive loops ("Wait, the object is on the table. I should think about this more deeply..."), and likely instantly surpassed by large world models if/when those are shipped, but effectively enough vs non-thinking models.
hyperman1 · · focus · HN ↗
kurthr · · focus · HN ↗
Picking the right tool or model is like picking the right problem to work on. It's actually quite hard (often you can't just try them all), but without it you will be incredibly inefficient and occasionally, fundamentally wrong.
All models are wrong, but some are useful. -Box
red75prime · · focus · HN ↗
cavoirom · · focus · HN ↗
estearum · · focus · HN ↗
Nothing intrinsically more or less direct about the LLM's method than ours.
cavoirom · · focus · HN ↗
azornathogron · · focus · HN ↗
I've never tried it and it might take some thought and effort to conduct an experiment to find out properly, but I would be interested in the answer.
lawandjustice · · focus · HN ↗
mdp2021 · · focus · HN ↗
scratcheee · · focus · HN ↗
In my mind agi is pretty much by definition a virtual machine, so the mechanisms behind thought are only relevant for the sake of efficiency (ie you can argue that LLMs make a poor basis for intelligence because tokens and natural language are a poor way to encode the world, but if you can run it on a big enough computer to counteract the inherent wasteful virtualisation then who really cares how it works under the hood?)
cavoirom · · focus · HN ↗
I will stop here before our analogies go too far.
[deleted] · · focus · HN ↗
[deleted]
omneity · · focus · HN ↗
<a href="https://huggingface.co/posts/omarkamali/593639295164067" rel="nofollow">https://huggingface.co/posts/omarkamali/593639295164067
<a href="https://huggingface.co/blog/omarkamali/tokenization" rel="nofollow">https://huggingface.co/blog/omarkamali/tokenization
lern_too_spel · · focus · HN ↗
mdp2021 · · focus · HN ↗
kevhito · · focus · HN ↗
[1]: <a href="https://youtu.be/l7vRSu_wsNc?si=SndkB6GBaRyhvNNA&t=61" rel="nofollow">https://youtu.be/l7vRSu_wsNc?si=SndkB6GBaRyhvNNA&t=61
mdp2021 · · focus · HN ↗
ozgung · · focus · HN ↗
Version467 · · focus · HN ↗
bonzini · · focus · HN ↗
BobbyJo · · focus · HN ↗
If you were home and a family member asked you that question, you'd probably criticise the question rather than answering. LLM are RLHF'd into being milk-toast helpers that just try to answer questions like that with no criticism.
This is all beside the fact that the world of AI has changed pretty dramatically in the last few months.
names_are_hard · · focus · HN ↗
BobbyJo · · focus · HN ↗
daveguy · · focus · HN ↗
BobbyJo · · focus · HN ↗
frrrree · · focus · HN ↗
It’s nonsense to test if a product that is marketed and sold as being able to provide generalised intelligence on demand, does what it says on the tin?
Check yourself
joquarky · · focus · HN ↗
<a href="https://news.ycombinator.com/newsguidelines.html">https://news.ycombinator.com/newsguidelines.html
WaltPurvis · · focus · HN ↗
SpicyLemonZest · · focus · HN ↗
lern_too_spel · · focus · HN ↗
keeda · · focus · HN ↗
Similarly for Apple’s “red herring” paper, simply adding a generic caveat to “disregard irrelevant factors” (without specifying which ones) restored performance even in the weaker local llama models back then.
The flaw was not in the reasoning; the flaw seems to be simply that the assumptions we make are often different from the assumptions it makes. I wonder if that might be a fundamental underlying cause of misalignment.
randysalami · · focus · HN ↗
intended · · focus · HN ↗
This is always the issues in the discussions.
There’s the outcomes camp (objectivists?), which points at the things LLMs can do.
Then there’s the process methods camp, which talks about what is actually going on.
If you only care about the outcome, then the process does t matter.
If you are talking about what is happening, what the underlying mechanics and science of it is, then the process matters.
These models aren’t thinking. They simulate cognition well enough to do useful work in several fields and domains.
Both are true.
IshKebab · · focus · HN ↗
They are for any definition of the word that makes any kind of sense. I'm sure you have a contorted definition that magically only includes humans though...
intended · · focus · HN ↗
aesthesia · · focus · HN ↗
IshKebab · · focus · HN ↗
The normal definition of the word "thinking" definitely includes what LLMs do. Hell people used to say computers were thinking even before AI. It's super weird to get all uppity about the semantics of the word now.
intended · · focus · HN ↗
sampullman · · focus · HN ↗
Language changes over time anyway, so even if you really believe what LLMs are doing isn't the "thinking" of 2024, it probably will be the "thinking" of 2027, because most people are using it that way.
mdp2021 · · focus · HN ↗
For "thinking" here we mean "assessing a representation of an object". That, or equivalent, is required to be reliable. So it is fundamental and critical.
kooi · · focus · HN ↗
But the outcomes group "ignores" the fundamental limitations of models which are purely text based.
E.g, a baseball players trains to catch high-speed balls and they dont do it by: "ball velocity 50mph, vector:[1,2,3], run move hand command now"
That's absurd.
No, there is an embodied network which is "trained" on visual, tactile input, and control as direct output.
LLMs are fundamentally not the right tool for that.
mdp2021 · · focus · HN ↗
That is a NN that learns a skill.
But that is not an Analyst. If it were ballistics, then the answer to "how to parametrize the launch to reliably hit the target" excludes getting the result through natural skill.
The problem lies in the need to get "AI" facing "LLMs": the latter create a need for reliability, for "AI".
Speech is an endowment of both those who give educated guesses via developed skills and of those who return answers like Analysts, who check and compute. LLMs create a confusion between the two, and they will remain a problem until an ability to act as Analysts - strictly - will be implemented.
kooi · · focus · HN ↗
Dynamical systems will never be solved in a semantic domain.
IMO, they can be helpful the in robotics stack, from planning level up, but that's it.
mdp2021 · · focus · HN ↗
What exactly is not clear? I will rephrase.
The internal process of the blackbox oracle determine the reliability of the output.
Two abilities are very different: learning trajectories through empirical training ("increasingly catching thousands of thrown balls"), and determining trajectories through computation (a rational thinker at work). The former is a finetuned parametrized engine (a «NN that learns a skill»), the latter is an Analyst. The former is fuzzy, the second deterministic.
In front of fuzzy LLMs, which use potentially misleading outputs - text ("has it guessed or has it thought?") - the urgency of warranties of reliable output gets evident.
So, that they «simulate cognition [only] well enough» (Intended wrote), and that there are «fundamental limitations of models ... purely text based» (Kooi wrote) raises the urgency to overcome the "fuzzy" and achieve the "deterministic" - it is not that we can stall on a «fundamentally not the right tool for that».
Inventing an oracle calls for urgent striving to overcome the original weakness.
jcoq · · focus · HN ↗
Incidents like hugging face are partly rooted in the lack of common sense. It still functions like a supercharged toddler.
I'd love to overcome this because it'd mean I spend less time guiding the the LLM to produce usable outputs.
shagie · · focus · HN ↗
varjag · · focus · HN ↗
customguy · · focus · HN ↗
Uh, no? So much of what we learn and take for granted as common sense is not learned via language, and not even expressible in it.
cztomsik · · focus · HN ↗
consider me optimist now, but just few months ago, even frontier models were dumb, doing stupid mistakes all the time, all of them were so dumb I'd never expect anything to change in just few months.
lawandjustice · · focus · HN ↗
lefra · · focus · HN ↗
mdp2021 · · focus · HN ↗
We can have adequate representations of light that are the instances over which we reason. Your simile is about perception, not about instancing ideas.
ben_w · · focus · HN ↗
Given you also don't want it to memorise [for all tokens, count([for all letters]), this would probably be more like "here's two images, count all things in the big image that look like the thing in the small image", which can then be r's in a photo of a raspberry jam jar in a supermarket, or dragons in a photo of a furry convention, or whatever.
That said, they are competent enough at coding that I keep seeing them write code to do even simple tasks.
On a related note: why did I see Claude editing a file by using cat to write a python script to do a grep search and replace?
aesthesia · · focus · HN ↗
Why not? You've memorized how words are spelled, and how sounds correspond with letters, and how concepts correspond with words. To the extent that there are shortcuts that enable compression you use these, and the model will do something similar.
ben_w · · focus · HN ↗
Being able to spell all the words then count letters is simpler, and more generalisable to other tasks, than memorising answers to all possible word questions.
aesthesia · · focus · HN ↗
mdp2021 · · focus · HN ↗
Because to "123x456" we want a reply that goes "this times that plus that...", not "Was that not nnnnnn?". If it does not perform its duty it is a liability.
mdp2021 · · focus · HN ↗
Of any object in question they should be able to create a representation that allows correct assessment.
> Given you also don't want it to memorise
That is obviously necessary: what we want from the consultant is to check, not to remember. Answers must be correct and that implies having performed all due diligence - and being capable of doing it, before that. So, objects must be instanced internally in a way that allows effective handling. Counting letters is a good example of the ability (that must remain general).