‹ BackHN Continuity

Thread

GLM-5.3 and the spread of advanced cyber capabilities

254 points · 241 comments · Philpax

  1. bitexploder · · focus · HN ↗
    This just makes me want a home lab capable of running GLM 5.3 at a 4bit quant.

    Also, for what it is worth Qwen Flash Next 3.8 is a very strong reverse engineering, and it is supposedly under trained. Qwen 3.8 27B is also strong. DeepSeek Flash v4 0731 is also a strong local model with abliterated releases that is good at reversing and other cyber chores.

    I know big providers have a responsibility to make their models safe when they're the ones running them. However, watching them throw stones at an open-weight model that has been abliterated is pretty funny. Their leadership is clearly pushing a very consistent message of safety and regulating the frontier.

    1. glimshe · · focus · HN ↗
      Are there local models that can run on 8-12GB GPUs that can help reverse engineer retro software (DOS games and applications)?
      1. bitexploder · · focus · HN ↗
        If you are really invested and have some system RAM you could get a 3-4 bit quant Qwen 3.5 35B-A3B running. There are builds that do expert caching, keeping the hot experts in cache. For something like disassembly, you're looking at being able to fit, if you have, say, 11 to 12 GB of VRAM, you could get at least three hot experts. For pure disassembly tests, I would say that would be pretty fast. A 4-bit quant is pretty decent and maintains most of the smarts of the larger quants. Depending on the GPU I would expect a decent token rate. It is medium strength local model, but if you harness and ground it well I expect it can reconstruct C code for you. The quality of your disassembler will matter here.

        If you have a lot of system RAM you could technically run Qwen Flash Next. On a 4080 with 16GB of RAM and 128GB of DDR5 I get ~35-40 t/s. And it is very capable.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.