How GLM built its own inference infrastructure
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
How GLM built its own inference infrastructure
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
dada216 · · focus · HN ↗
axiosgunnar · · focus · HN ↗
[dead]
freakynit · · focus · HN ↗
gpugreg · · focus · HN ↗
dzonga · · focus · HN ↗
china has cheap abundant power, now they can make their own inference chips (which was supposed to be a chokepoint), their models yeah can be 6 months behind the frontier - but most people don't need frontier models - small models r more than enough.
my only wish was labs like Mistral would make their own inference chips or partner up eg with established / new chip makers or companies like Oxide.
vatsachak · · focus · HN ↗
g023 · · focus · HN ↗
jchook · · focus · HN ↗
Only the US sees this as a competition, new space race, Cold War, etc.
shostack · · focus · HN ↗
MrBuddyCasino · · focus · HN ↗
Also, don’t underestimate data retention and such. Big Corp will never send their LLM traffic to China.
throwawayqqq11 · · focus · HN ↗
cbg0 · · focus · HN ↗
spacebanana7 · · focus · HN ↗
Perhaps more than the price advantage is the prioritisation in grid infrastructure. As a strategic industrial concern in China you almost certainly get easy access to transformers, grid connections, water etc which is a big bottleneck in the US.
embedding-shape · · focus · HN ↗
They must have hit really hard scaling limits if the prices were hiked so much so quickly.
broodbucket · · focus · HN ↗
lompad · · focus · HN ↗
And you can bet GLM is still ridiculously subsidized, just not as ridiculously as Anthropic and OpenAI.
chobbledotcom · · focus · HN ↗
jdiff · · focus · HN ↗
breakingcups · · focus · HN ↗
pyrophane · · focus · HN ↗
bbor · · focus · HN ↗
For [API usage](<a href="https://openrouter.ai/z-ai/glm-5.3-flash#providers" rel="nofollow">https://openrouter.ai/z-ai/glm-5.3-flash#providers) they charge a bit more than the very cheapest providers of GLM-5.3-Flash, but not so much that a big price difference would make sense.
Daviey · · focus · HN ↗
disiplus · · focus · HN ↗
world2vec · · focus · HN ↗
Can I ask where are you using all those tokens?
wartywhoa23 · · focus · HN ↗
p2detar · · focus · HN ↗
absqueued · · focus · HN ↗
tokai · · focus · HN ↗
world2vec · · focus · HN ↗
disiplus · · focus · HN ↗
world2vec · · focus · HN ↗
embedding-shape · · focus · HN ↗
disiplus · · focus · HN ↗
_0ffh · · focus · HN ↗
rubslopes · · focus · HN ↗
buckle8017 · · focus · HN ↗
Daviey · · focus · HN ↗
embedding-shape · · focus · HN ↗
Daviey · · focus · HN ↗
I now exclusively use <a href="https://omp.sh/" rel="nofollow">https://omp.sh/ as my harness:
I set it up so it never works in the main branch so subagents etc don't step on each others toes, and only merges back when complete: <a href="https://github.com/Daviey/mario/blob/main/.omp/hooks/pre/worktree-guard.ts" rel="nofollow">https://github.com/Daviey/mario/blob/main/.omp/hooks/pre/wor...
A good AGENTS.md is essential: <a href="https://github.com/Daviey/mario/blob/main/AGENTS.md" rel="nofollow">https://github.com/Daviey/mario/blob/main/AGENTS.md
I then provide specifications for what I want, making sure it is unit tested.
asp_hornet · · focus · HN ↗
<a href="https://docs.z.ai/legal-agreement/privacy-policy" rel="nofollow">https://docs.z.ai/legal-agreement/privacy-policy
andy_ppp · · focus · HN ↗
asp_hornet · · focus · HN ↗
criley2 · · focus · HN ↗
Also why Meta gets a +1, just charge less money on the training path.
orf · · focus · HN ↗
asp_hornet · · focus · HN ↗
To be fair, none of us are sure of anything and I think that’s the part that’s most irritating
orf · · focus · HN ↗
asp_hornet · · focus · HN ↗
orf · · focus · HN ↗
orf · · focus · HN ↗
andy_ppp · · focus · HN ↗
Havoc · · focus · HN ↗
Very good - but I'm on a legacy plan. And coming up on a renewal that would put me on the watered down current plan. But with 50% legacy discount think it may be worthwhile. If I go to a competitor I'd be paying market rate.
>They must have hit really hard scaling limits if the prices were hiked so much so quickly.
Not really scaling - their plans were initially comically subsidized even more so than what the western providers are doing. More advert for an upstart than commercially priced.
probst · · focus · HN ↗
_aavaa_ · · focus · HN ↗
The max plan will provide ~1,100 USD of GLM-5.3 or ~260 USD of GLM-5.3-flash per month for 168 USD. I can personally attest to these numbers through omp (~97% cache hit rate).
Unless you are able to highly parallelize (your work, you won't be able to hit your hourly or weekly quota using the flash model simply because it's so slow.
They give you ~3x more flash tokens, which maybe comes out to ~2x more actual work after accounting for the extra thinking it does to achieve the same result. The mental model, for not getting angry, is 5.3 is fast mode by default, and you can disable fast mode for 2x the work output at 1/3-1/10th the speed.
They're serving me 5.3 at ~40 tok/s and 5.3-flash at 30 tok/s (according to omp).
Schlagbohrer · · focus · HN ↗
That is shocking. Is it per-token I wonder?
workbreak · · focus · HN ↗
_aavaa_ · · focus · HN ↗
bleonard · · focus · HN ↗
Just checking now: recent runs tau3[1] was at 96% and toolathlon[2] was at 90%
[1] <a href="https://www.induction.ai/docs/benchmarks/tau3" rel="nofollow">https://www.induction.ai/docs/benchmarks/tau3 [2] <a href="https://www.induction.ai/docs/benchmarks/toolathlon" rel="nofollow">https://www.induction.ai/docs/benchmarks/toolathlon
bbor · · focus · HN ↗
jensb1 · · focus · HN ↗
bingud · · focus · HN ↗
drbscl · · focus · HN ↗
bbor · · focus · HN ↗
In the PRC, they[1] leaked tons of national secrets on the PRC's latest AI campaigns, the inner workings of their "opinion monitoring" (read: performative panopticon) and "stability" (read: violent oppression) departments, Chengdu's whole CCTV network, direct-energy weapons plans, espionage activities in Syria to hunt down Uyghur refugees, and god knows what else that Anthropic didn't divulge to us common folk.
In the US, it's very clearly an attempt to rip off a competitor. I'm not sure how else you could possibly see it. Even if you're a distillation fan in general (which A. why and B. plz don't), they did this through a network of Japanese and Signaporean shell accounts, presumably at least some of which were abusing Anthropic's subscription service in a ToS double-whammy, as it would be exorbitantly expensive otherwise. They also had to hack around Anthropic's API to get CoT traces, which seems impossible to explain away as anything innocent.
I've been beating the "China isn't necessarily an enemy, it's gonna take us all to handle AI" drum for literally years, but this attack was just... gross. Gross in scale and gross in arrogance. Not a good sign for the dawning alignment crisis, to say the least :(
TL;DR: Use these services if you want, but know that you're supporting aggressive escalations and companies that very clearly don't give a flying fuck about violating the law, much less your ToS. So... buyer beware, I guess.
[1]: For clarity, Z.ai was not alone in this, nor were they most egregious attack -- Moonshot.ai (kimi) took that coveted prize. DeepSeek was involved, too.
Bluestein · · focus · HN ↗
jensb1 · · focus · HN ↗
[dead]
dgellow · · focus · HN ↗
> use these services if you want, but know that you're supporting aggressive escalations and companies that very clearly don't give a flying fuck about violating the law, much less your ToS
From my European point of view the same risk/concerns apply when using US providers
cmrdporcupine · · focus · HN ↗
bbor · · focus · HN ↗
jLaForest · · focus · HN ↗
podocarp · · focus · HN ↗
pjc50 · · focus · HN ↗
Alignment is meaningless; as you've noticed, humans aren't all that "morally aligned".
If the tool needs safety measures it should be kept in a safe enclosure like we do with CNC machines, furnaces, and so on.
bbor · · focus · HN ↗
- Is what [DICTATOR/MURDERER/CRIMINAL] bad, or merely not to your taste? If the latter, then you have no coherent reason to argue they should be punished. We would never imprison people who don't like vanilla ice cream because 51% of the population does like it.
- If another culture had a deeply held belief to [HORRIBLE_THING] to, say, children, would you just shrug and say "different strokes for different folks"? What if [MURDERER] just had a different culture?
- No, the fact that nature is red in tooth and claw does not disprove morality; we are very, very, very far from our pre-rational, animalistic roots, and to go back now would be unthinkable.
Thousands of scientists have been studying this problem for 76 years now, going on 77; your hunch about physical machines does not overrule their findings about the capabilities and tendencies of minds wrought from sand. I think this is just blatantly false, likely based in a misunderstanding of criminal law vs. civil law. Civil courts still deal with legality.The broader discussion of why distillation is bad and dangerous and immoral is left as an exercise for the reader, as it was above with the parenthetical. It's not a complex argument; I guarantee you understand it if you're reading this.
The same thing as above -- the fact that laypeople can not think of a criminal charge that they've heard on Law & Order that corresponds to this behavior does not mean that it's legal. It's textbook fraud, regardless of what particular detail you focus on. I think the fact that it's happening at such a large scale is proof that very smart, well-resourced labs in China (the producers of the world's best OS models, including the incredible GLM-5.3-Flash) think it's feasible. I'm not sure it's productive to question them in the absense of any indication to the contrary.This is a great question still, not trying to shut you down. But I think the fundamental issue is a misunderstanding of what distillation is -- it's not directly stealing literal atomic parameters and piling them up. They might try to focus on substructures within these massive networks, but even that isn't strictly necessary for a distillation attack.
Sorry, I never linked it! This is from the latest Anthropic safety report (of "Anthropic Houtis build missile" fame), and no, these cannot be hallucinated -- the leaked secrets were inputs, not ouputs. <a href="https://www.anthropic.com/threat-intelligence-report-september-2026" rel="nofollow">https://www.anthropic.com/threat-intelligence-report-septemb... This is just blatant word games, sorry. I'm sure intended in good faith, and I understand the impulse -- I consider myself a radical anti-IP slacktivist, after all. But "both things involve information transfer" is just not a coherent point; lots of things fit that description. These are cyberattacks. Yes, anyone can cyberattack cyberattackers. But, y'know... an eye for an eye... I am not at all concerned with the value of the resulting artifacts as assessed by the (already totally unhinged) NYSE et. al. I am concerned about user respect, law following, truth telling, blatant cyber warfare at a time of rising tensions, accidental data leakages at a scale that'd be hard to fathom 5 years ago, bad-faith public postures, and a general distaste for fraud.ux266478 · · focus · HN ↗
The struggle you have with pinning down moral relativism I think betrays the fact that your understanding of it is low quality. A good ear mark is, can you name a single passage from a treatise on moral relativism that you like? If you can't find a single aspect of a philosophical construction to advocate for, that means you don't actually understand it.
You also keep sliding between conflating moral relativism with moral anti-realism and even moral nihilism at one point. These are different axes.
> Realism + Absolutism
All moral facts converge for every single person. As a matter of fact, they aren't facts at all. Morals are a hinge of the Wittgenstein variety.
> Realism + Relativism
Moral facts are derived from frameworks, or contexts, which are themselves grounded facets of reality.
> Anti-Realism + Absolutism
Kantian Constructivism. There are no facts which are coherent without a framework, but moral claims are still necessarily only capable of being universal.
> Anti-Realism + Relativism
Moral reality is constructed by the framework, grounded only to the framework.
tuesdaynight · · focus · HN ↗
lelanthran · · focus · HN ↗
I mean, if they get to distill other's IP, why can't others distill their IP?
whizzter · · focus · HN ↗
The Antrophic article mentions "16 million" conversations, GLM models are in the 700-300 billion parameter ranges and while the frontier sizes aren't know but Gemini suggests Astra and Mythos are at around 10 trillion. That'd amount to extracting 40k parameters per conversation without a lot of errors if it was just a distillation (from an unknown source/algorithm as opposed to distilling your own model).
Now, I can imagine these conversations being used as a verification step that they're not missing stuff in their training, and that their models are capable of most of the same things, but that's mostly confirming that they've stolen the same data from the public as Antrophic/OpenAI has stolen already.
Or am I missing something here that makes real "distillation" feasible?
mitxela · · focus · HN ↗
Laurel1234 · · focus · HN ↗
[dead]
woadwarrior01 · · focus · HN ↗
<a href="https://x.com/EricSimons/status/2099252922098061714" rel="nofollow">https://x.com/EricSimons/status/2099252922098061714
phoghed · · focus · HN ↗
butterNaN · · focus · HN ↗
pjc50 · · focus · HN ↗
No real reason to respect any terms they might want to impose. Besides, if you want to break TOS, just have an agent do it; "everyone" running these things agrees there's no corporate or moral liability for what your AI does.
_aavaa_ · · focus · HN ↗
lelanthran · · focus · HN ↗
HarHarVeryFunny · · focus · HN ↗
I wonder how you imagine that China built their own space station? Reliant on using American made duct tape, perhaps?
Do you realize how reasoning models are being trained nowadays? You design/build simulation environments to run agents in, with the environment providing the RLVR "verification" scoring. So why won't Ziphu use GLM to build their own RL training environments? Do you think they are not doing this?
mitxela · · focus · HN ↗
tefkah · · focus · HN ↗
Statements dreamed up by the utterly deranged.
binsquare · · focus · HN ↗
sixeyes · · focus · HN ↗
xyzsparetimexyz · · focus · HN ↗
koe123 · · focus · HN ↗
phoghed · · focus · HN ↗
jdiff · · focus · HN ↗
phoghed · · focus · HN ↗
> Do you have anything that proves this one way or another that isn't based on vibes or shoddy benchmarks?
They clearly aren’t talking about RSI here, but that model development has stalled in general.
jdiff · · focus · HN ↗
koe123 · · focus · HN ↗
xyzsparetimexyz · · focus · HN ↗
pizza234 · · focus · HN ↗
Ironically, many benchmarks being maxxed out, and quite quickly, so new ones have to be created.
simonw_simonw_ · · focus · HN ↗
[dead]
georgefloydd · · focus · HN ↗
[dead]
imp0cat · · focus · HN ↗
ttoinou · · focus · HN ↗
glimshe · · focus · HN ↗
heresiarch39 · · focus · HN ↗
scarmig · · focus · HN ↗
etamponi · · focus · HN ↗
red75prime · · focus · HN ↗
TL;DR Buckmaster and Alpöge haven't solved Navier-Stokes blow up.
How information can get so distorted when it's trivial to fact check?
scarmig · · focus · HN ↗
How, exactly, does "it was Claude that solved Navier Stokes, not ChatGPT!" get you to "AI has hit a wall and stalled"? That's, not to put too fine a point on it, incoherent, and is just noise thrown into the discussion to avoid grappling with the fact that AI continues to rapidly improve.
owebmaster · · focus · HN ↗
meyer3423 · · focus · HN ↗
[dead]
ramon156 · · focus · HN ↗
simonw_simonw_ · · focus · HN ↗
[dead]
azan_ · · focus · HN ↗
fc417fc802 · · focus · HN ↗
simonw_simonw_ · · focus · HN ↗
[dead]
tefkah · · focus · HN ↗
pizza234 · · focus · HN ↗
Some call this "The singularity" (e.g. Hinton).
This is actually a core danger postulated by the, let's call it, "worrying" scenario - see AI 2027 (to be clear, I think its timeline is not realistic).
> Statements dreamed up by the utterly deranged.
Evidently, and tragically, it will take catastrophes to show that deranged are the ones deriding the worried crowd.
itsalwaysgood · · focus · HN ↗
Everything you think you are is just what you can imagine in a single moment. Nothing more, nothing less.
OhNoNotAgain_99 · · focus · HN ↗
[dead]
rob74 · · focus · HN ↗
Honestly, I have no idea what z.ai is either (I'm aware of an AI-enabled editor called Zed, but that's under zed.dev), so it's a bit presumptuous from them to assume that everyone is familiar with their product...
fxwin · · focus · HN ↗
Also I feel like the obvious way to read the very first sentence is that GLM is a language model
> As we develop GLM, the model sometimes exhibits capabilities that surprise us
jbonatakis · · focus · HN ↗
ma2kx · · focus · HN ↗
drbscl · · focus · HN ↗
Come on now
Mashimo · · focus · HN ↗
Where GLM-5.3-Flash is the newest "small / fast" model.
0x457 · · focus · HN ↗
peri-cl · · focus · HN ↗
<a href="https://artificialanalysis.ai/#intelligence-category-tabs" rel="nofollow">https://artificialanalysis.ai/#intelligence-category-tabs
bogdan · · focus · HN ↗
rob74 · · focus · HN ↗
tokai · · focus · HN ↗
bogdan · · focus · HN ↗
Appreciate the clarification. For me it was the "F" in "WTF" that tipped me. Other than that, it's more than fair for you to not know what GLM is. Things are moving so fast that I would be surprised if anyone can keep track of it all. Cheers, have a grand day!
HarHarVeryFunny · · focus · HN ↗
Why would you be reading their corporate blog posts if you don't even know who they are?!
Argonautlabs · · focus · HN ↗
[dead]
tipsytoad · · focus · HN ↗
Argonautlabs · · focus · HN ↗
[dead]
zozbot234 · · focus · HN ↗
Havoc · · focus · HN ↗
GLM has in the past been more technical rather than speculation about future development on RSI etc.
Also curious whether those 100k accelerators are entirely locally made. If that's genuinely end to end on all components including lithography, memory, design etc then that is quite a feat.
dude250711 · · focus · HN ↗
Schlagbohrer · · focus · HN ↗
MaKey · · focus · HN ↗
ElectricalUnion · · focus · HN ↗
Schlagbohrer · · focus · HN ↗
bitexploder · · focus · HN ↗
Catloafdev · · focus · HN ↗
Also, Deepseek V4 Flash can be run relatively well in hybrid 2-bit quantization on 128gb devices, with way better results than you'd expect for a typical 2-bit quant.
Those are currently the 'smartest' options for that memory level.
alightsoul · · focus · HN ↗
ElectricalUnion · · focus · HN ↗
alightsoul · · focus · HN ↗
CamperBob2 · · focus · HN ↗
Fable will still dump you back into Opus at the drop of a hat, but at least it will tell you when it happens.
HarHarVeryFunny · · focus · HN ↗
Just like the rest of the world, including the US (Intel, Micron), SMIC are currently using ASML lithography equipment (DUV, not EUV), but Shanghai Aishengna are now moving into early production with their own DUV machines, with SMIC and CXMT as early customers.
There is also a state sponsored Chinese EUV development underway.
wiz21c · · focus · HN ↗
almaight · · focus · HN ↗
[dead]
jonstewart · · focus · HN ↗
HarHarVeryFunny · · focus · HN ↗
All I can recall reading from OpenAI about what they have actually done in the name of "RSI" is using one of their models to help automate the training process.
cmrdporcupine · · focus · HN ↗
That's more than Anthropic has done though.
HarHarVeryFunny · · focus · HN ↗
Ziphu seem much more matter of fact about it. To me it' a shame that they've decided to the use this "RSI" name, but at least they are being fairly specific about what they mean by it, while the western companies seem to want to invite you to think it's more than just dogfooding and automation.
_aavaa_ · · focus · HN ↗
[dead]
zicohacks · · focus · HN ↗
HarHarVeryFunny · · focus · HN ↗
In addition to Huawei who make the Ascend series that Ziphu are using, there are also at least a half dozen or so other Chinese companies also making their own AI accelerators.
0xbadcafebee · · focus · HN ↗
ipsod · · focus · HN ↗
Seems to work for them.
freakynit · · focus · HN ↗
Almost everyone knew that these sanctions would backfire within a few years. You can't really put sanctions that have noticeable negative effects on bigger economies. They only work for small to medium economies. I believe sanctions on any economy in top 10 would fail.
verdverm · · focus · HN ↗
They are likely more concerned a out AMD taking market share, and I suspect that geopolitics will leave them with largely non overlpping customer bases.
overfeed · · focus · HN ↗
That's not what's happening in the auto-industry, Chinese EVs are easily outselling American ones. Why would GPUs be any different?
verdverm · · focus · HN ↗
This is another example where we might ask "why would it be any different?"
HarHarVeryFunny · · focus · HN ↗
It'd be ironic if Europe ended up banning ASML (a Dutch company) from shipping to US instead!
verdverm · · focus · HN ↗
Mark Carney said earlier today at the EU, "the goal is not self-sufficiency, but collective resilience." It is through collective action that they can find more resilience to our American Antics.
PorciiVorbesc · · focus · HN ↗
ASML's Dutch EUV machines depend on importing US made EUV light sources from California to build them, and depends on sales of their machines to the US to stay afloat since they're not exactly flying off the shelves domestically.
So maybe the EU is not that stupid and suicidal to cripple their economy further by cutting off one of their biggest pay-piggies for no reason.
In fact, as an anecdote to your hypothetically scenario, the US is the one with the most leverage since if they want, they can ban the export of EUV light sources to ASML and sell them instead to Canon and Nikon, sinking ASML and propping up Japan's semi industry instead, if relations sour and they see Japan as a warmer ally than the EU. Hypothetical.
HarHarVeryFunny · · focus · HN ↗
PorciiVorbesc · · focus · HN ↗
It's on US soil and subject to US laws. In case of conflict that's all that matters.
HarHarVeryFunny · · focus · HN ↗
PorciiVorbesc · · focus · HN ↗
aurareturn · · focus · HN ↗
So until China solves the ASML problem, there won't be any flooding.
HarHarVeryFunny · · focus · HN ↗
A bit late for that now that they are moving into early production with their own.
Catloafdev · · focus · HN ↗
Which they are in progress on: <a href="https://www.reuters.com/world/china/how-china-built-its-manhattan-project-rival-west-ai-chips-2025-12-17/" rel="nofollow">https://www.reuters.com/world/china/how-china-built-its-manh...
aurareturn · · focus · HN ↗
HarHarVeryFunny · · focus · HN ↗
China have been really squeezing all they can out of DUV machines, but I'm not sure how much node size really matters for AI competition - more of a cost issue (more chips/power for same FLOPs) than anything, and TSMC & NVIDIA's healthy profit margins need to be considered too.
aurareturn · · focus · HN ↗
What is the timeline like? 1 year? 5 years? 10 years?
HarHarVeryFunny · · focus · HN ↗
Small numbers perhaps, but this is happening right now.
What will the numbers be in 5-10 years time? Who knows, but ASML took about 5 years to go from 20/yr to 100+/yr.
Of course politically the world may well be quite different in 5 years time, as may be the AI market.
aurareturn · · focus · HN ↗
We also don't know how small of nm chips they can manufacture. If it's 20nm, it's practically useless for advanced AI chips.
HarHarVeryFunny · · focus · HN ↗
Now, the US is left out in the cold with little influence left, themselves now the ones with an anti-missile shortage.
ffsm8 · · focus · HN ↗
You seem to misunderstand what Protectionism is. This is not an example of it not working. If anything, it is any example of it working. Because Protectionism is about protecting your industry from foreign competition - exactly what China decided to do.
pjc50 · · focus · HN ↗
Brazil is one to watch if they manage to achieve political stability, as is Nigeria if it can transition from being a petrostate.
0xbadcafebee · · focus · HN ↗
my-huge-pony · · focus · HN ↗
smallmancontrov · · focus · HN ↗
AuthAuth · · focus · HN ↗
mitxela · · focus · HN ↗
pjc50 · · focus · HN ↗
skybrian · · focus · HN ↗
menaerus · · focus · HN ↗
> Compared with our initial baseline on the same hardware, we achieved a 3× improvement in end-to-end serving performance, reaching hardware efficiency and per-token cost comparable to mainstream NVIDIA GPUs. This demonstrates that Chinese chips can support frontier-model inference efficiently and economically at scale.
verdverm · · focus · HN ↗
a34729t · · focus · HN ↗
gpt5 · · focus · HN ↗
The other part is that it’s a bit of a meme here to say that the chip restriction is actually helping China (or shall I say, coordinated effort?). For once, we know that China has put a lot of pressure on the US to relax these controls multiple times. In addition to large chip smuggling networks (e.g. 22% of NVidia’s worldwide revenue magically comes from Singapore, and the ratio has been growing).
Lastly, assuming acceleration in AI (which we ARE seeing), there might not be time to China to catch up. The best estimate right now is that the first EUV chips from China will not come out before 2030. By that time who knows how powerful AI will be.
All I’m saying is that the discussion is so one sided and a bit baselesss with no nuance, that it seems either a meme/groupthink in the community or coordinated. If anything, the data suggests that the US should increase its export controls and better track the tech supply chain if it wants to further curb Chinese progress.
miroljub · · focus · HN ↗
Or, to put it bluntly, you are a proponent of the trade war.
edgyquant · · focus · HN ↗
dominotw · · focus · HN ↗
never heard of this 'talking point' . who is even saying this?
stanfordkid · · focus · HN ↗
vcryan · · focus · HN ↗
dominotw · · focus · HN ↗
but your supposed corollary was
> China wasn’t working hard to build their own chips
kind of dishonest ?
true_religion · · focus · HN ↗
The unnatural and sudden nature of the US chip embargo has galvanized the Chinese companies to speed up development. It makes economic sense to do so rather than pay inflated rates to essentially smuggle a necessary good, while never knowing if the embargo will be tightened and those under-the-table supply chains dry up as well.
Yes, there was a plan, but what was planned for 5 years from now is being done today, because if you wait 5 years, you'll be so far behind thanks to the effect of the embargo.
edgyquant · · focus · HN ↗
menaerus · · focus · HN ↗
gpt5 · · focus · HN ↗
menaerus · · focus · HN ↗
gpt5 · · focus · HN ↗
bigbadfeline · · focus · HN ↗
[dead]
menaerus · · focus · HN ↗
pianopatrick · · focus · HN ↗
Because if the AI is just really good at writing software, I am not sure why China has to "catch up". Seems to me China can just treat AI like other technologies. Let the US pay most of the R&D costs, then come after and treat the technology as a commodity. Sell it better and cheaper and at scale.
I am not sure why it's a big deal for China if China is a few months behind the US in terms of the very frontier AI. Just like I'm not sure if it's a big deal which country has the biggest super computer in the world.
gpt5 · · focus · HN ↗
I am not inventing a winner take it all scenario, that is what driving the race.
pianopatrick · · focus · HN ↗
If AI cannot stop nuclear ballistic missiles, then how does AI give the US power over China? Seems to me that after AI we have the same mutually assured destruction we have today. In which case America cannot really stop China from progressing or threaten China too much.
Seems to me being super human intelligent does not inherently guarantee power. Just like the smartest person in the world does not always win in a fist fight or a gun fight or war. I.e. Stephen Hawking might have been smart, but he did not have physical strength and power. Power and smarts are not always related.
gpt5 · · focus · HN ↗
pianopatrick · · focus · HN ↗
gpt5 · · focus · HN ↗
klrefg · · focus · HN ↗
gpt5 · · focus · HN ↗
rcxdude · · focus · HN ↗
gpt5 · · focus · HN ↗
And we already see AI’s super human ability in computer systems. It will be able to be Omni present and completely control every computer in any place, and bribe and social engineer and mimic any person digitally. All while doing in an instant what a team of humans would require years.
I’m really surprised by the lack of logical thinking in HN when it comes to China.
deterministic · · focus · HN ↗
HSO · · focus · HN ↗
ok, show it
i`ve recorded a few people who saw this coming in 2022 and i can tell it wasnt many
menaerus · · focus · HN ↗
goodmythical · · focus · HN ↗
conmod278 · · focus · HN ↗
Tadpole9181 · · focus · HN ↗
Weren't they just saying that it's self-evident that preventing a manufacturing superpower - one with a significant pool of industrial engineers and effectively unlimited money & government backing - from acquiring some good is a temporary measure? Because if they have sufficient incentive, they would just... Learn to build it themselves?
After all, their entire nation is built around building things, and catching up / leap-frogging is much easier than starting from scratch.
HSO · · focus · HN ↗
in 2022, mainstream and westoid virtualism opinion was ``china is cut off from the future now`` (cue the clownface here). ``we cut off chinas legs`` yadayada
there were very few informed people who differed.
now after the fact every dumb little shmuck ``always knew`` ``it was obvious``
really? then show the record where you predicted it. in 2022.
not now, pretending to be wise and informed :)))
it`s the same like in palestine. in 2023, 24 so many people pretended to look away or not see. now everybody has always been against it all of a sudden.
fucking bozos.
browningstreet · · focus · HN ↗
maybe where you were paying attention.
but lots of people who saw what they were doing with renewable energy and EVs rightly suggested this would put them on the path of independence w/r/t GUI and LLMs.
this isn't like the ol' days where "japan can't do software". we knew china can.
dvngnt_ · · focus · HN ↗
swyx · · focus · HN ↗
HSO · · focus · HN ↗
and also, arguably jensen`s track record is such that i wouldnt exactly lump him with the mainstreamers )))
chvid · · focus · HN ↗
ux266478 · · focus · HN ↗
[1] - <a href="https://www.eastwestcenter.org/sites/default/files/private/iegwp001_1.pdf#:~:text=move%20from%20catching-up%20to%20forging-ahead%20in%20semiconductors" rel="nofollow">https://www.eastwestcenter.org/sites/default/files/private/i...
[2] - <a href="https://www.usitc.gov/publications/332/journals/chinese_semiconductor_industrial_policy_prospects_for_success_jice_aug_2019.pdf#:~:text=China's%20gap%20with%20leading%20international%20semiconductor%20firms%20has%20consistently%20narrowed%20over%20time" rel="nofollow">https://www.usitc.gov/publications/332/journals/chinese_semi...
HSO · · focus · HN ↗
[dead]
stogot · · focus · HN ↗
NorthSouthNorth · · focus · HN ↗
verdverm · · focus · HN ↗
edgyquant · · focus · HN ↗
aatd86 · · focus · HN ↗
aurareturn · · focus · HN ↗
Winners: Huawei, SMIC, CXMT,Chinese ASML-competitors, OpenAI, Anthropic, Amazon, Microsoft, Google, Meta.
Losers: Chinese AI labs, Nvidia, AMD, TSMC, Micron, SK Hynix, Samsung, Intel
throwaw12 · · focus · HN ↗
In the short term maybe yes, in the long term, maybe they are the winners, they can build on top of cheap inference stack and eventually win on pricing
aurareturn · · focus · HN ↗
bigmadshoe · · focus · HN ↗
klrefg · · focus · HN ↗
bigmadshoe · · focus · HN ↗
Catloafdev · · focus · HN ↗
I don't think your information is entirely accurate.
aurareturn · · focus · HN ↗
Just logic.
Catloafdev · · focus · HN ↗
aurareturn · · focus · HN ↗
Intel being up 100% has nothing to do with China market being restricted. They could be up 200% instead if the China market is free.
Catloafdev · · focus · HN ↗
klrefg · · focus · HN ↗
mullingitover · · focus · HN ↗
PorciiVorbesc · · focus · HN ↗
Nokia shares were also up after the iPhone got announced. Shares don't really mean much in rational terms, just vibes and speculation of clueless masses.
conorcleary · · focus · HN ↗
abtinf · · focus · HN ↗
You can’t really hurt a country that has a culture with a positive attitude toward growth.
christina97 · · focus · HN ↗
Half-arsed export restrictions are the best of both worlds for these firms: enough of an incentive to take homegrown hardware seriously, yet not aggressive enough to cause meaningful handicap in the meantime.
brazukadev · · focus · HN ↗
What's the alternative other than a military/naval embargo?
chvid · · focus · HN ↗
rising-sky · · focus · HN ↗
If someone is capable of doing something, and your goal is to prevent them for doing it, the worst thing you can do is to make it necessary for them to do it.
thih9 · · focus · HN ↗
Just leave that damned prophesied hero alone. Don’t banish them, don’t attempt to kill them while they’re young, or send them on an impossible quest. Your entitled meddling is exactly what puts them on their path. Some level of boring coexistence may have been an option.
rising-sky · · focus · HN ↗
thih9 · · focus · HN ↗
Also note that grandparent comment is about both folklore and current events.
aesthesia · · focus · HN ↗
rising-sky · · focus · HN ↗
[dead]
unrented7977 · · focus · HN ↗
jeffybefffy519 · · focus · HN ↗
Slightly similar is US sanctions on NK & Russia causing a restriction of foreign currency - so they turned to cyber crime to get it...
ux266478 · · focus · HN ↗
thehappypm · · focus · HN ↗
chung8123 · · focus · HN ↗
tokai · · focus · HN ↗
gpugreg · · focus · HN ↗
Bawoosette · · focus · HN ↗
menaerus · · focus · HN ↗
GLM: 80 USD (pro), 168 USD (max) -> with "limited-time event" discount this becomes 56 USD and 117.6 USD
I also don't understand why are they so much costlier, and I would also like to give it a try.
ipsod · · focus · HN ↗
GLM's "Max" plan is (was?) equivalent to 3x Claude's 20x ($200) plan.
Bawoosette · · focus · HN ↗
Anthropic's Pro is $20 and corresponds to Z.ai's Lite at $18
Anthropic's 5x Max is $100 and corresponds to Z.ai's Pro at $80
Anthropic's 20x Max is $200 and corresponds to Z.ai's Max at $168
reacharavindh · · focus · HN ↗
I have both plans. Claude monthly €20 and Z’s €18 monthly. Running GLM-5.3 high on their monthly plan will hit quotas absurdly fast compared to Opus 5 High on Claude code. It’s almost unusable for AI driven development. I ended up using the Z plan for using GLM-5.3 as a detailed security reviewer and adversarial feedback. For that, it is much better than Opus which will flag and bail out for even simple security tasks that are aimed at defense.
SSLy · · focus · HN ↗
reacharavindh · · focus · HN ↗
But, it was enough for a customer like me who tried them at good faith to walk away and find their competitors..
I like the diversity of LLMs as of today and prefer to not tie myself to one big plan with any vendor. If they don’t prefer me as a customer, then I will accept that, and move away.
SSLy · · focus · HN ↗
alexjplant · · focus · HN ↗
SSLy · · focus · HN ↗
auspiv · · focus · HN ↗
throwa356262 · · focus · HN ↗
0xbadcafebee · · focus · HN ↗
wolttam · · focus · HN ↗
esseph · · focus · HN ↗
kamranjon · · focus · HN ↗
dilyevsky · · focus · HN ↗
HarHarVeryFunny · · focus · HN ↗
If you look at how many years the whole NVIDIA and CUDA ecosystem has been evolving, it's certainly impressive how they've just stood up and optimized this CUDA-free 100,000 node cluster in just a few months.
saagarjha · · focus · HN ↗
KronisLV · · focus · HN ↗
konart · · focus · HN ↗
And at the same time you have pretty strict limits to your usage, so in many cases you can't even let it work all night, as you will reach your limit faster than that.
yorwba · · focus · HN ↗
9cb14c1ec0 · · focus · HN ↗
a012 · · focus · HN ↗
Because they don’t have to. Most of the time money would buy you newest and/or more hardwares so there’s low/minimal interest to optimize the code or approach.
Art9681 · · focus · HN ↗
alex_duf · · focus · HN ↗
vblanco · · focus · HN ↗
kingstnap · · focus · HN ↗
US labs are quite cut throat about dealing with stuff costing them money (inference). This sort of engineering excellence doesn't always feel that way because they are simultaneously quite lax about stuff costing other people money.
esafak · · focus · HN ↗
Signed, a customer.
ElectricalUnion · · focus · HN ↗
chrisjj · · focus · HN ↗
bguberfain · · focus · HN ↗
sigbottle · · focus · HN ↗
- How do you develop new tests and metrics to capture the distinctions you want? Do you work in industries where the abstractions have mostly stabilized? Or is the work in coming up with the measurements themselves?
- How do you stable-ly solve the causality issue? You can prove for one commit in time that some variable was causing an issue by changing it - fine. But that may not address deeper design issues that may not be expressed by that one implementation, if that makes sense. It seems like there's always a meta-level you can go to; sometimes justifiably, sometimes unjustifiably. I liken this to solving individual memory leaks - that's "causal", you can prove that yes, this line of code was the cause - but the deeper issue may be that you're using a memory unsafe language in the first place, and in some sense that's "outside" your ontology. Sure, in the case of memory safety, we've thought about this for decades, so in some sense that's stable enough where it's outside, but known; but what happens when it's outside, but unknown?
I suppose at some point this delves into, "What is good software engineering" in general. And that's not to mention the integration and legacy effects - maybe there's some "fundamentally new better way" to do something, but that requires changing the entire product and a ton of constraints.
And if the response to this is, "Aha! That's the hard part about software engineering!", then how can I get experience in this kind of thing? I know it's abstract but at my current company I haven't developed anything that's lasted more than half a year. I'm wondering if at some point I'd be better of trying to bootstrap experience off of open source.
furyofantares · · focus · HN ↗
kgeist · · focus · HN ↗
gpugreg · · focus · HN ↗
jchook · · focus · HN ↗
The way these “AI is too powerful now” articles read about Mythos, Fable, GLM, etc is completely incongruent with my experience using them. It feels like they are all trying to position themselves to influence government policy.
pixl97 · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
[deleted] · · focus · HN ↗
[deleted]
krttherealest · · focus · HN ↗
a11r · · focus · HN ↗