No, whether or not it's intentional, maybe can be debated. But there's definitely an experience of a model losing horsepower quickly after launch.
The claim was that there's an experience of a model losing power. Your claim amounts to "No, you are not experiencing what you say". That's quite a claim for you to make with no data and no argument.
Has there ever been any measurement of this, of any sort? Honest question. I frequently see a plural of anecdotes to that effect, but I've not seen a concrete statement of fact or measurement that could be scrutinized or tested in any way.
If so, please share. This should be measurable, and I'm glad this project is measuring it.
Answers in the form of additional anecdotes, stated with even greater passion but still lacking a statement that could be tested and falsified, would validate my exact concern.
Official tweet from Tibo from OpenAI confirming they ran experiments tweaking “juice” (reasoning effort mapping values) and have reverted them: <a href="https://x.com/thsottiaux/status/2076495156757577895" rel="nofollow">https://x.com/thsottiaux/status/2076495156757577895
As he confirmed, for at least a period of time, and for some users, “Sol xhigh” was actually “Sol high”, etc.
I was just wondering if, like certain processors, bugs get fixed and the speed goes down. Like, they find it's doing things it shouldn't, restrict it, and harm the throughput.
Yeh it's absurd that people claim this all the time. It's some crazy conspiracy theory and when you ask for examples nothing ever shows up.
It would be economical suicide from anthropic and OpenAI to actually need models intentionally.
But hey I guess it's hard with technology that truly seems like magic.
People say if you'd bring electricity to the middle ages you'd be called a witch and burned. The same is happening to the model labs here because they are bringing tech that the world isn't ready for yet.
Dishonest representation. The links provided do not show this at all, but very clearly that it has been isolated incidents that people extrapolate into false evidence.
solenoid0937 · · focus · HN ↗
solfox · · focus · HN ↗
nba456_ · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
omani · · focus · HN ↗
nba456_ · · focus · HN ↗
AnimalMuppet · · focus · HN ↗
solenoid0937 · · focus · HN ↗
voiceeh · · focus · HN ↗
nba456_ · · focus · HN ↗
doginasuit · · focus · HN ↗
redanddead · · focus · HN ↗
Kiro · · focus · HN ↗
nimchimpsky · · focus · HN ↗
[dead]
Craighead · · focus · HN ↗
bradfa · · focus · HN ↗
jyoung8607 · · focus · HN ↗
If so, please share. This should be measurable, and I'm glad this project is measuring it.
Answers in the form of additional anecdotes, stated with even greater passion but still lacking a statement that could be tested and falsified, would validate my exact concern.
dannyw · · focus · HN ↗
As he confirmed, for at least a period of time, and for some users, “Sol xhigh” was actually “Sol high”, etc.
dude250711 · · focus · HN ↗
empath75 · · focus · HN ↗
wccrawford · · focus · HN ↗
raincole · · focus · HN ↗
jascha_eng · · focus · HN ↗
It would be economical suicide from anthropic and OpenAI to actually need models intentionally.
But hey I guess it's hard with technology that truly seems like magic. People say if you'd bring electricity to the middle ages you'd be called a witch and burned. The same is happening to the model labs here because they are bringing tech that the world isn't ready for yet.
sumedh · · focus · HN ↗
Links have been provided by others in this post.
Kiro · · focus · HN ↗
sumedh · · focus · HN ↗
Ant denied them at first though.