‹ BackHN Continuity

Thread

GLM-5.3 and the spread of advanced cyber capabilities

254 points · 241 comments · Philpax

  1. CharlieDigital · · focus · HN ↗
    Anthropic has to use this wedge (and future ones) to move regulatory action against the Chinese models or their IPO is going to be really problematic.

    (Ironic, though, that I haven't heard of any Chinese models "escaping" which Anthropic and OpenAI both seem to have issues with...)

    Like Chinese electric cars, the American producers cannot compete without regulatory action. Yes, I understand that the Chinese government this and that in both the automotive and AI industries.

    But reality is what it is as a consumer: it's a cheaper product that's almost as good or better in some cases. And in the case of these open weight models: I can run it on my own infra and not give any data to anyone.

    1. enraged_camel · · focus · HN ↗
      >> And in the case of these open weight models: I can run it on my own infra and not give any data to anyone.

      It's worth noting that the overwhelming majority of people who use Chinese models don't do this. Yes, it is nice to have the option, and there are US-based inference providers that claim to not send your data to China and maybe indeed don't, but in the grand scheme of things, we need to remember the adage that became popular during the social media era: if something is free (or, in this case, close to free), you are the product.

      1. jacquesm · · focus · HN ↗
        I actually do do this. I'm not sure who the 'overwhelming majority' is and where you got the data (link would be appreciated) but everybody that I know that runs these is doing so on their own infra.
        1. enraged_camel · · focus · HN ↗
          Context is useful. The parent said: "it's a cheaper product that's almost as good or better in some cases"

          The only open models that are "almost as good or better in some cases" require massive amounts of RAM. I posit that most people cannot afford a decked out Mac Studio, and therefore run the smaller "flash" variants on more normal devices. The issue is that those are nowhere near frontier-level in terms of capability.

          1. CharlieDigital · · focus · HN ↗
            "Running your own infra" also includes managed infra like Bedrock, Foundry, etc.

            Not just your local machines.

            Enterprises are where you see this adoption. Legal, finance, tax; sensitive context where the data must be contractually opaque to external parties.

            1. jacquesm · · focus · HN ↗
              Precisely. I have figured out a nice recipe that is quite affordable, 288G of VRAM for a little under 20K, it takes some fiddling though, but once it works it is really neat.

              PCIe is incredibly powerful tech.

            2. enraged_camel · · focus · HN ↗
              >> "Running your own infra" also includes managed infra like Bedrock, Foundry, etc.

              Managed infra is, by its very definition, not your own infra. It's infrastructure someone else sets up and manages for you.

              1. CharlieDigital · · focus · HN ↗
                It's your own infra because there's a distinct line item cost for it versus OpenAI.

                Same way you say "my apartment" and not "my landlord's apartment". It's your place while you're renting it.

                1. enraged_camel · · focus · HN ↗
                  Sorry, that distinction makes zero sense. Either way you're paying a monthly opex cost, compared to it being your own hardware, in which case it would be a fixed capital expense. Which is what "your own infra" means. I worked in managed IT services for 15 years. Trust me, the terminology is important and words don't suddenly start to mean what you want them to mean.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.