The OpenAI Decisions API needs a confidence you can trust
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
The OpenAI Decisions API needs a confidence you can trust
Unofficial Hacker News client; not affiliated with Y Combinator.
bob1029 · · focus · HN ↗
I am currently using the Responses API with my clients, which is mandatory to get at non-zero reasoning effort in the latest models. Luna without reasoning turned on might as well be a model from early 2025. This is not how anyone is using this. Responses with 5.6-luna+ and high+ reasoning level feels pretty close to the Star Trek computer experience for me.
Attempting to recreate the OAI reasoning model capabilities at home seems like a pointless quest now. You will never get the access into the base models that the frontier companies have internally. You will also never have access to an engineering team with that kind of capacity. You must submit to the black box if you want the advertised performance figures.
lxgr · · focus · HN ↗
Importantly, it's probably also not what it's been trained to do.
Until quite recently, OpenAI used to ship dedicated "instant" and "reasoning" models. Newer ones seem to have reasoning levers that can be turned down all the way to zero, but that doesn't mean they don't take a significant performance hit when doing that.
bananaflag · · focus · HN ↗