OpenAI GPT–6 Astra breaks Enigma message that has resisted solution since 2005
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI GPT–6 Astra breaks Enigma message that has resisted solution since 2005
Unofficial Hacker News client; not affiliated with Y Combinator.
jtrn · · focus · HN ↗
Stubborn for a long time because the message used a completely different key from the rest of that day's traffic. Everyone assumed it shared the daily key. The original transcription had errors. The left rotor turned over at letter 72, which is rare and breaks standard crib attacks.
What is cool, if true, is that it was a 2 day collab between the Leffer and Astra. To me this shows the importance of human in the loop, was still all also showing how immensely power of llm tools. But I think it’s getting a bit silly how much anrticles ignores the driving force (the person) in breakthroughs like this.
exfalso · · focus · HN ↗
jtmarl1n · · focus · HN ↗
Macuyiko · · focus · HN ↗
"However, the most astonishing thing about this break is that the GPT–6 Astra did it entirely on its own. Carter Leffer only directed GPT–6 Astra to see if it could break any of the unbroken Enigma messages published on the Crypto Cellar Research web page."
I mean... I'm all for collaboration but I think this case is pretty clear, no?
alerighi · · focus · HN ↗
We are fooling to me, there is no intelligence in these models, they just apply methods that were invented by humans without any consciousness on what they are doing.
allturtles · · focus · HN ↗
huty · · focus · HN ↗
At the margin the innovative human matters.
What that means for the rest of society is TBD.
TeMPOraL · · focus · HN ↗
huty2 · · focus · HN ↗
redsocksfan45 · · focus · HN ↗
[dead]
jstanley · · focus · HN ↗
Technically you built it yourself and the builder was just a minor collaborator?
[deleted] · · focus · HN ↗
[deleted]
cicko · · focus · HN ↗
jfyi · · focus · HN ↗
true_religion · · focus · HN ↗
Hearing a guy built his home in a week with the power of nails and a hammer would have been novel in era of mortise and tenon.
I actually found an article about raving about how fast nail production was thanks to machining advances in 1790 and that it would bring great value: <a href="https://digital.libraries.psu.edu/digital/collection/pabooknews/id/1054/" rel="nofollow">https://digital.libraries.psu.edu/digital/collection/pabookn...
It’s seems to me humans haven’t changed, just which machines we praise.
radium3d · · focus · HN ↗
Now we just code in english and the computer does the rest.
platevoltage · · focus · HN ↗
radium3d · · focus · HN ↗
It is not yet the same as C++ or Python, for two reasons:
1) Ambiguity is still the default. Formal languages force you to resolve it up front.
English lets you paper over it until the model or the compiler (the human) notices.
2) The "compiler" (the LLM) is statistical and non-deterministic. Same prompt, different day, different bugs. A real language has a spec.
The practical move is to treat English as a high-level specification language, keep the generated artifacts inspectable, and "still know enough of the lower layers to notice when the translation went wrong."^1
[1] This is the key that is where humans can still be necessary, or at least another pass through the LLMs to decide on the best path, in the compiled code. Compilers for other languages do the same C -> Binary, etc.
A conventional compiler is bound by an as-if rule. It can take many internal routes, but the observable behavior has to match the language spec. Same source, same defined semantics. If two gcc runs emit different binaries, the program is still supposed to compute the same answers on the same inputs. That is why people treat the source as the artifact and the binary as disposable.
An LLM compiling English has no as-if rule unless you add one. "Sort the users by last active" can become a stable sort, an unstable sort, a SQL order by, an in-memory timsort, or a query that drops people with null timestamps. All of those can look like success. They are different programs. The model is not optimizing under a spec. It is filling in the parts you did not write.
A human who can read the destination language still notices when the chosen path is the wrong program.
A second model pass can compare paths, but only if you give it a way to score them: tests, types, invariant
So the historical analogy still holds, with one correction. JavaScript and Python were dismissed for being too easy, but they already had grammars and evaluators. English is easier still, and the evaluator is a statistical translator that will invent a dialect if you let it. The practical move stays the same: treat English as the spec language, pin the generated artifacts behind tests, and keep enough fluency in the lower layer to see when the translation chose a different program than the one you meant. The human is not required because the computer is weak. The human is required because the source language still leaves room for more than one destination.
platevoltage · · focus · HN ↗
English is not code just because a technology was developed that could make educated guesses based on being trained with other code that people have written as to what the code generated should look like.
Also, much of this reply looks LLM generated.
speed_spread · · focus · HN ↗
TeMPOraL · · focus · HN ↗
speed_spread · · focus · HN ↗
NetMageSCW · · focus · HN ↗
speed_spread · · focus · HN ↗
It's a machine - when we attribute it human concepts such as victory, mistakes, ethics, we also shift responsibility away from the operator and soon enough bad people will use it as an umbrella. The agent ate my homework! The agent bombed a school! Bad agent, no!
1attice · · focus · HN ↗
I think this sounds somewhat less silly in English because "automotive" and "automative" don't have the same hyper-visible affinity, but all the same; you may want to consider the car as a more viable analogand.
UpsideDownRide · · focus · HN ↗
The shortcomings should really be obvious by now to anyone honest. And the marketing distortion being oushed out is just tiresome and detrimental for all of us.
anthonyrstevens · · focus · HN ↗
shuvrojit · · focus · HN ↗
TeMPOraL · · focus · HN ↗
Even when the report literally says the LLM did it on its own?
Let's not over-correct in the direction of knowing better than the first party.
jfyi · · focus · HN ↗
Not mention it also says this...
> We are still analysing the GPT–6 Astra logs to see exactly how it executed the break.
[deleted] · · focus · HN ↗
[deleted]
ricksunny · · focus · HN ↗
For everything else, there’s Astracard
0c3ca83z · · focus · HN ↗
[dead]
dev_tty01 · · focus · HN ↗
"After analysing the unbroken messages on the website, it decided that the most promising message was Nr. 172, MVUEH and it also quickly suspected that the plaintext of Nr. 173, SIPVX ..."
dyauspitr · · focus · HN ↗
PunchyHamster · · focus · HN ↗
WithinReason · · focus · HN ↗
"However, the most astonishing thing about this break is that the GPT–6 Astra did it entirely on its own."
pixelesque · · focus · HN ↗
"Carter Leffer only directed GPT–6 Astra to see if it could break any of the unbroken Enigma messages published on the Crypto Cellar Research web page."
[deleted] · · focus · HN ↗
[deleted]
gre · · focus · HN ↗
awesome! keep going
great work! keep going
ec109685 · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
TeMPOraL · · focus · HN ↗
vorticalbox · · focus · HN ↗
eru · · focus · HN ↗
weiran · · focus · HN ↗
antii · · focus · HN ↗
andriy_koval · · focus · HN ↗
nonethewiser · · focus · HN ↗
chrisjj · · focus · HN ↗
That's not correct for the content.
"However, the most astonishing thing about this break is that the GPT–6 Astra did it entirely on its own. Carter Leffer only directed GPT–6 Astra to see if it could break any of the unbroken Enigma messages published on the Crypto Cellar Research web page."
[deleted] · · focus · HN ↗
[deleted]
mrcwinn · · focus · HN ↗
josu · · focus · HN ↗
Today, all it takes to get to the top 3 is "/goal get to the top of the leaderboard".
The human-in-the-loop is only a temporary measure until the models get good enough.
zahlman · · focus · HN ↗
wolfi1 · · focus · HN ↗
serbuvlad · · focus · HN ↗
Just like with Kasparov's Centaur Chess, the idea of a 'human in the loop' is just a necessity due to current limitations.
There will hopefully (?) come a time one day when human beings provide only ultimate value judgments, and everything else is done by machines. Or it may not.
But I don't think betting your ego on the idea that you will be useful in the loop for very long is very wise.
skybrian · · focus · HN ↗
ctoth · · focus · HN ↗
This? Still? After everything?
Buddy, you're living in the future. In a science fiction novel. Please get used to it.
skybrian · · focus · HN ↗
(Also, getting people to think about the future rather than the present is a classic con. Looking at an empty field: "can't you just see the potential here?")
pyrale · · focus · HN ↗
The novel title: "Don’t Build The Torment Nexus".
sifar · · focus · HN ↗
Unfortunately, there are enough people in the world who think that is the future everyone should live in and are actively working to bring it about.
And so, one must adapt. .
myrmidon · · focus · HN ↗
I find this suprising kinda. If I got a choice right now, there's a bunch of dystopian Scifi that I would instantly go for (out of sheer curiosity), e.g. the Murderbot universe.
Are there fantasy worlds that you would want to live in? I feel this is a bit of a suspect benchmark in the first place because books typically want some kind of tension/conflict which you won't get if everyone is just gratefully living their best life.
anthonyrstevens · · focus · HN ↗
serbuvlad · · focus · HN ↗
I do expect centaurs to outperform other systems for many types of tasks for years to come (and am kind of betting on this to keep getting paid).
But what I'm talking about is ego. It your ego is tied up with (a) your intelligence or (b) your ability to perform task X; you will probably be humbled this century.
goatlover · · focus · HN ↗
One is humanist, the other is anti-human, (in the end goal at least).
krisoft · · focus · HN ↗
Where do you see the commenter say this?
goatlover · · focus · HN ↗
In their previous grandparent post. They do reserve a place for human value judgement, but I doubt even that remains if everything else has been automated. The machines will decide what we value. We already see that to some degree with algorithmic engagement and targeted advertisements.
serbuvlad · · focus · HN ↗
dyauspitr · · focus · HN ↗
banannaise · · focus · HN ↗
myrmidon · · focus · HN ↗
Compare with go (boardgame): Centaur go was basically not ever a thing.
Other applications behaved kinda similarly (AI Starcraft/Dota/...), where we had decent "human-like" heuristics from the get go and the Centaur concept could never really shine, much less for a decade or more.
I'd also like to stress that past progress in this mainly happened for the love of the game, while the (economical) incentives to replace human office workers are... high.
zahlman · · focus · HN ↗
What you describe doesn't actually sound to me like the worthy goal it might initially come across as. Struggle helps us feel alive.
1attice · · focus · HN ↗
Therefore Astra could also have done this comment better
mannyv · · focus · HN ↗
Don't people actually read anymore?
HDThoreaun · · focus · HN ↗
moffkalast · · focus · HN ↗
jfyi · · focus · HN ↗
edit: I think that's going to be my go to on "you aren't an artist" from now on. "No! I'm an AI researcher!"
jtrn · · focus · HN ↗
I found that Leffen even said the explanatory website took about 99 times more effort than the codebreaking itself. And he said that he set the direction and pushed, and the model did the execution. How much steering "pushed forward" involved is not disclosed anywhere, but in this instance, it seems to be more a case of "Human pointed at hard task and AI did an awesome job mostly by itself." Tho how much he was a simple meat-ralph-loop is not entirely clear.