If the sigmoid approacheth, the first one to get a good model into silicon, at acceptable benchmarks and silicon rejection %, will make some truckloads. Most people don't need SOTA for most of their problems.
the TAM might be a real number, but the addressee might already be walking around, such as any qwen3.8 model. Qwen3.8-Flash-Next is churning out some fascinating capabilities as we speak.
There's only so many math problems you can solve without ROI.
I used to think that, but I've started to come around to the fact that better AI comes with progressive unlocks that create new use cases. I suspect in 18 months, one of the primary consumers of tokens will be people using LLMs for one-shotting increasingly complicated games rather than just white collar work.
Then there's the matter of engineering and synthesizing consumer objects completely personalized for users!
The ecosystem around LLMs is becoming just as or more important than the model. Having fast access to a web crawl and other reference material is important.
LarsDu88 · · focus · HN ↗
I was also shocked by how quickly AMD acquired Talaas. AMD may be preparing for the next way (ultra fast inference, and embodied AI inference)
cyanydeez · · focus · HN ↗
the TAM might be a real number, but the addressee might already be walking around, such as any qwen3.8 model. Qwen3.8-Flash-Next is churning out some fascinating capabilities as we speak.
There's only so many math problems you can solve without ROI.
LarsDu88 · · focus · HN ↗
Then there's the matter of engineering and synthesizing consumer objects completely personalized for users!
treis · · focus · HN ↗