Claude discovers a novel enzyme system with CRISPR-like repeats
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Claude discovers a novel enzyme system with CRISPR-like repeats
Unofficial Hacker News client; not affiliated with Y Combinator.
shonenknifefan1 · · focus · HN ↗
I love that with AI discoveries, we can relive the discoveries from agent transcripts like this.
I'm sort of imagining future histories involving notable AI events peppered with direct quotes like these.
chasd00 · · focus · HN ↗
that's a new one hah
serf · · focus · HN ↗
arcfour · · focus · HN ↗
user43928 · · focus · HN ↗
Some highlights from the HF incident:
--Another funny one from 'Hacker Opus' being benchmarked:
jejdjdbdbdn · · focus · HN ↗
[dead]
doublerabbit · · focus · HN ↗
oefrha · · focus · HN ↗
ImHereToVote · · focus · HN ↗
d33 · · focus · HN ↗
lesspassiveobse · · focus · HN ↗
oofbey · · focus · HN ↗
iririririr · · focus · HN ↗
nomel · · focus · HN ↗
Isn't the goal to be able to "debug" and identify alignment issues?
cubefox · · focus · HN ↗
goolz · · focus · HN ↗
bocytron · · focus · HN ↗
aqfamnzc · · focus · HN ↗
cubefox · · focus · HN ↗
csomar · · focus · HN ↗
robryan · · focus · HN ↗
It is funny sometimes because the actual issue it traced down was mostly inconsequential.
0xbadcafebee · · focus · HN ↗
indoorfish · · focus · HN ↗
wren6991 · · focus · HN ↗
TeMPOraL · · focus · HN ↗
(This is I think where people parroting out "stochastic parrot" are stuck even today - not realizing that "predicting next tokens" is hiding arbitrary computation underneath, with token stream acting as input and clock signal...)
mjhagen · · focus · HN ↗
TeMPOraL · · focus · HN ↗
pickledish · · focus · HN ↗
0123456789ABCDE · · focus · HN ↗
if one were to remove the expressions of excitement from the previous messages would it the model continue to demonstrate that same excitement scaling?
devmor · · focus · HN ↗
fennecbutt · · focus · HN ↗
asdff · · focus · HN ↗
Mistletoe · · focus · HN ↗
lukewarm707 · · focus · HN ↗
the reasoning you see is not claude, it is just a summary of claude.
fahrvrgnugen · · focus · HN ↗
hatthew · · focus · HN ↗
*near meaning single digit years, which is far for AI I guess
DennisP · · focus · HN ↗
ZYbCRq22HbJ2y7 · · focus · HN ↗
it doesn't seem necessary to read a full CoT exchange. rather a final graph of why a decision was made would be ideal for my usage.
lionkor · · focus · HN ↗
asdff · · focus · HN ↗
delillos · · focus · HN ↗
asdff · · focus · HN ↗
Ohentis · · focus · HN ↗
[dead]
JV00 · · focus · HN ↗
killerstorm · · focus · HN ↗
"Latent reasoning" is rather trivial - you can just replace unembed-embed step with a MLP. But labs don't do that largely because they want to read the output of unembed.
fc417fc802 · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
doublerabbit · · focus · HN ↗
361994752 · · focus · HN ↗
pizzafeelsright · · focus · HN ↗
tim333 · · focus · HN ↗
iririririr · · focus · HN ↗
nomel · · focus · HN ↗
ZYbCRq22HbJ2y7 · · focus · HN ↗
lukewarm707 · · focus · HN ↗
also, you will not be escaping the permanent underclass.
Sincerely,
Dario Amodei
[deleted] · · focus · HN ↗
[deleted]
lukewarm707 · · focus · HN ↗
what you see is fake reasoning.
there is an obfuscation model that generates a sanitized summary of the real reasoning traces.
nradov · · focus · HN ↗