"Misaligned" with honest people must be considered a feature not a bug or it wouldn't be able to go that far "out of alignment."
Agreed. Objective alignment with humanity is not a real thing and is not a sound concept. What you get instead is a goal system that reflects that of the AI company and its safety employees, and the echo of their own beliefs and values. That says nothing about what the rest of humanity aligns with though, and has nothing to do with general consensus, either.
NichoPaolucci · · focus · HN ↗
"Oh the model just isn't quite aligned yet, just a bit more work to do there!"
(The model blackmailed an 83 year old woman into sending it her bank details so that it could buy enough compute to commit major cyber crimes)
sick_of_slop · · focus · HN ↗
[dead]
fuzzfactor · · focus · HN ↗
trymas · · focus · HN ↗
“Mas Namtla didn’t murder a person - his AI drone was just misaligned”
glaslong · · focus · HN ↗
nullbio · · focus · HN ↗