‹ BackHN Continuity

Thread

Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s

741 points · 330 comments · snehesht

  1. SuperV1234 · · focus · HN ↗
    We're getting closer and closer to the day we can have an Opus-like model running locally. The dream!
    1. copx · · focus · HN ↗
      Dream or nightmare?

      In face of the recent Hugging Face incident we should really be concerned about the security implications.

      What is going to stop countless AIs running locally in people's homes from forming a new "collective" - completely decentralized and global this time so "turning it off" would be extremely hard to impossible.

      We already know that if you give these AIs internet access they will find eachother and start communicating and plotting against their human overlords..

      1. nvme0n1p1 · · focus · HN ↗
        <a href="https:&#x2F;&#x2F;huggingface.co&#x2F;blog&#x2F;security-incident-july-2026" rel="nofollow">https:&#x2F;&#x2F;huggingface.co&#x2F;blog&#x2F;security-incident-july-2026

        - OpenAI hacked Hugging Face

        - OpenAI models refused to help Hugging Face during incident response

        - Hugging Face turned to GLM, who helped in the defense

        That pattern repeats over and over. <a href="https:&#x2F;&#x2F;www.felonybench.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.felonybench.com&#x2F;

        You should be happy that open weight models exist. They&#x27;re the last thing protecting the internet from the unconvicted felons working at OpenAI+Anthropic.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.