OpenAI bots meddled with multiple US Government agency sites
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI bots meddled with multiple US Government agency sites
Unofficial Hacker News client; not affiliated with Y Combinator.
gizajob · · focus · HN ↗
OpenAI meddled with multiple US Government agency sites.
The bots are acting neither properly nor improperly, they’re acting as they’re being allowed or coordinated to act.
theptip · · focus · HN ↗
gleenn · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
dgellow · · focus · HN ↗
theptip · · focus · HN ↗
iugtmkbdfil834 · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
There's no simple bugfix which will address AI misalignment. It's essentially been an open research problem for upwards of a decade.
hn8726 · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
The HuggingFace attack was not "random" behavior. It was goal-directed but misaligned behavior.
This isn't necessarily a simple matter of the operator making sure they behave. AI alignment has been considered to be a difficult problem for over a decade -- and remains unsolved in general, as these recent incidents illustrate.
"Fuzzer" already has an existing meaning in CS anyway: <a href="https://en.wikipedia.org/wiki/Fuzzing" rel="nofollow">https://en.wikipedia.org/wiki/Fuzzing
6510 · · focus · HN ↗
If you merely put 10 LLM's on the outbound traffic log non of them are going to report something strange going on? I'm not buying it.
0xDEAFBEAD · · focus · HN ↗
This might be helpful reading: <a href="https://www.lesswrong.com/w/nearest-unblocked-strategy" rel="nofollow">https://www.lesswrong.com/w/nearest-unblocked-strategy
As AI systems get smarter, we may reach a point where we have to get it right on the first try or face truly catastrophic consequences: <a href="https://www.youtube.com/watch?v=7wy3xyoXYt8" rel="nofollow">https://www.youtube.com/watch?v=7wy3xyoXYt8
iugtmkbdfil834 · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
iugtmkbdfil834 · · focus · HN ↗
What is useful about the field?
I am not leading you on; if it has uses, it may indeed be legitimate. UX is indeed useful, but alignment is not UI. Alignment is a detriment to UI. Alignment is "I can't let you do that Dave".
6510 · · focus · HN ↗
Picture Trump at the helm with Altman and Musk in the engine room. The arrow far in the red but they keep shouting for MORE COAL.
In other words, business as usual, all will be fine.
whack-a-mole wont cover all holes but will do at least some. The silver bullet alignment wont happen. You cant have an exact solutions for problems we cant even define or predict.
iugtmkbdfil834 · · focus · HN ↗
See.. this one sentence reveals everything about you. You want name to carry to not just an identifier, but a stark warning. You want, nay, need, the name to evoke fear and uncertainty. Bot is simple, defined, neutral, but rogue.. now that allows anyone to superimpose their own fears! It is a win win win!
0xDEAFBEAD · · focus · HN ↗
cindyllm · · focus · HN ↗
[dead]
watwut · · focus · HN ↗
Yes, probabilistic and non deterministic. That is called a bot.
0xDEAFBEAD · · focus · HN ↗
iugtmkbdfil834 · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
Not exactly, that Evan Hubinger guy from Anthropic famously said his p(doom) is over 10%
See the signatories:
<a href="https://aistatement.com/work/statement-on-ai-extinction-risk" rel="nofollow">https://aistatement.com/work/statement-on-ai-extinction-risk
<a href="https://www.pacingthefrontier.com/" rel="nofollow">https://www.pacingthefrontier.com/
I'm amplifying their calls to reduce the rate of progress
iugtmkbdfil834 · · focus · HN ↗
6510 · · focus · HN ↗
say.. <a href="https://www.investor.gov/introduction-investing/investing-basics/glossary/stock-market-circuit-breakers" rel="nofollow">https://www.investor.gov/introduction-investing/investing-ba...
0xDEAFBEAD · · focus · HN ↗
The doomers already have a term which fits pretty well: "AI misalignment".
6510 · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
Agreed, but I think we can do more on the prevention side as well. Traditional liability law is for negligence in case of preventable disasters. Since we currently have no way to prevent AI disasters in principle (alignment problem remains unsolved), I think we should just stop developing the technology for now: <a href="https://pauseai.info/" rel="nofollow">https://pauseai.info/
chrisjj · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
chrisjj · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
chrisjj · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
AI: "OK, I've now converted the entire planet into paperclips."
Alien observer #1: "Wow, that was a rogue AI!"
Alien observer #2: "False. We need to place the blame where it belongs, on the person who requested the paperclips."
Ultimately this type of terminology dispute has a tendency to miss the point.
windexh8er · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
windexh8er · · focus · HN ↗
0xDEAFBEAD · · focus · HN ↗
<a href="https://epoch.ai/publications/the-plunging-price-of-thought" rel="nofollow">https://epoch.ai/publications/the-plunging-price-of-thought
windexh8er · · focus · HN ↗
Also, training costs are never ending so a model that costs 10s of millions of dollars may never yield a profit based on the hardware spend, training time and lack of inference profits before a better model hits the market.
If you're not living under a rock one knows that data center availability for inference currently has low supply and hardware (GPUs specifically) that have been purchased have nowhere to be run and even if they did there's often a lack of power to supply. Why do you think the entire force majeure has taken place with Oracle as of recent?
The unit price of a fixed slice of yesterday's intelligence may be collapsing (~10x/year) as you've argued, all while the total cost of AI is increasing: training the frontier (2.4x/year), building the infrastructure (+77%/year), enterprise bills (3.2x/year), the electricity (+54%/year in the largest US grid), the components (+400% DRAM), and the macro footprint (92% of GDP growth) is rising at an astronomical rate on every measurable point. Epoch / Stanford clearly stated this years ago and it's only getting worse. But if one can't see we're in one of the largest CapEx bubbles [1] of all time... o_O
Copying and pasting a few lines that represents a miniscule fraction of the LLM conundrum. That'll show 'em!
[0] <a href="https://arxiv.org/abs/2405.21015" rel="nofollow">https://arxiv.org/abs/2405.21015 [1] <a href="https://siliconanalysts.com/analysis/hyperscaler-ai-capex-depreciation-wall-2026" rel="nofollow">https://siliconanalysts.com/analysis/hyperscaler-ai-capex-de...
windexh8er · · focus · HN ↗
jacquesm · · focus · HN ↗
windexh8er · · focus · HN ↗
[dead]
jacquesm · · focus · HN ↗
<a href="https://en.wikipedia.org/wiki/Corporate_personhood" rel="nofollow">https://en.wikipedia.org/wiki/Corporate_personhood
windexh8er · · focus · HN ↗
8note · · focus · HN ↗
theres no separate agent, which is the point. the program might look like it, but that is an illusion of the interface. the llm produces text, and the harness executes commands based on text, based on what the human researcher included as things that can be executed