If the sigmoid approacheth, the first one to get a good model into silicon, at acceptable benchmarks and silicon rejection %, will make some truckloads. Most people don't need SOTA for most of their problems.
the TAM might be a real number, but the addressee might already be walking around, such as any qwen3.8 model. Qwen3.8-Flash-Next is churning out some fascinating capabilities as we speak.
There's only so many math problems you can solve without ROI.
I used to think that, but I've started to come around to the fact that better AI comes with progressive unlocks that create new use cases. I suspect in 18 months, one of the primary consumers of tokens will be people using LLMs for one-shotting increasingly complicated games rather than just white collar work.
Then there's the matter of engineering and synthesizing consumer objects completely personalized for users!
I think it's true in the general case that people who want to create stuff (eg games) will create stuff and people who don't won't.
its hard bc everyone here is some form of builder. but there are people out there who just have almost no desire to create.
and even if we
lower the bar for creating to "hey llm make me a new video game to play tonight" that will still come with all the baggage of creation and turn folks off.
The ecosystem around LLMs is becoming just as or more important than the model. Having fast access to a web crawl and other reference material is important.
Other embodied labs are less gamer-y/consumer-y. WorldLabs didn't know whether to be a consumer tool, a creative tooling company, a robotics research lab.
Investors who have seen their materials told me this was their direct feeling.
AMD has acquired "AI stuff" for several billions at this point, yet somehow, doing "AI stuff" on their GPU/NPU stack still seems to be troublesome. Well, that is what I read from others anyway. Inference with llama.cpp on AMD GPUs pretty much just works.
I bet soon enough we can just ask AI to port stuff from CUDA to whatever language AMD is using. This is how nVidia will become a victim of their own success.
Keep going... extrapolate out further. There's a conclusion you beley will be true, but haven't fully thought through why that conclusion must be true if CUDA is no longer the moat it once was.
Who says that? I’m having a blast with rocm powering my R9700. Been able to run all models on day 0 at comparable speeds to Nvidia with similar memory bandwidth. Software is no longer the main bottleneck for AMD.
just use torch/vllm/sglang/llama.cpp and the rest of the ecosystem, most have decent support for AMD, it's not that different from CUDA once you get it working (and it is much easier now than it used to)
Really, it's just the Python frameworks on AMD consumer GPUs that suck. AMD datacenter GPUs seem to work with them well enough, and llama.cpp crushes it for the rest of us.
LarsDu88 · · focus · HN ↗
I was also shocked by how quickly AMD acquired Talaas. AMD may be preparing for the next way (ultra fast inference, and embodied AI inference)
cyanydeez · · focus · HN ↗
the TAM might be a real number, but the addressee might already be walking around, such as any qwen3.8 model. Qwen3.8-Flash-Next is churning out some fascinating capabilities as we speak.
There's only so many math problems you can solve without ROI.
LarsDu88 · · focus · HN ↗
Then there's the matter of engineering and synthesizing consumer objects completely personalized for users!
cyanydeez · · focus · HN ↗
Even if you buy singularity the people who needand afford it is dwindling.
This also assumes societal cohesion isnt destroyed by sed people.
dgently7 · · focus · HN ↗
its hard bc everyone here is some form of builder. but there are people out there who just have almost no desire to create.
and even if we lower the bar for creating to "hey llm make me a new video game to play tonight" that will still come with all the baggage of creation and turn folks off.
treis · · focus · HN ↗
echelon · · focus · HN ↗
Investors who have seen their materials told me this was their direct feeling.
ahartmetz · · focus · HN ↗
amelius · · focus · HN ↗
wmf · · focus · HN ↗
fragmede · · focus · HN ↗
nwah1 · · focus · HN ↗
fragmede · · focus · HN ↗
tybit · · focus · HN ↗
fragmede · · focus · HN ↗
homosapien97 · · focus · HN ↗
vlovich123 · · focus · HN ↗
alightsoul · · focus · HN ↗
sheepscreek · · focus · HN ↗
__rito__ · · focus · HN ↗
I don't know anything.
Things like training a Vision/RL/LM, downloading and running LMs, etc.
imjonse · · focus · HN ↗
whateverboat · · focus · HN ↗
ahartmetz · · focus · HN ↗
MrDrMcCoy · · focus · HN ↗