‹ BackHN Continuity

Thread

Claude Opus 5.5

1806 points · 1134 comments · km144

  1. ApolloFortyNine · · focus · HN ↗
    >Because Opus 5.5 is comparable to Claude Mythos 5.1 in biology and cybersecurity, we’re deploying it with safeguards similar to those on Claude Fable 5.1. Vetted organizations can apply today to our Life Sciences Verification Program to use Opus 5.5 for biology research. In the coming weeks we will also be expanding access to our Cyber Verification Program, and verified cybersecurity practitioners will be able to use Opus 5.5 for their work.

    Ah, they're spreading their limits to all their models it seems. Definitely not a good thing long term in my opinion.

    1. prettyblocks · · focus · HN ↗
      They're pushing their customers to their own competition by doing this.
    2. searine · · focus · HN ↗
      Great. Claude is basically useless for bioinformatics now.
    3. blfr · · focus · HN ↗
      Fable 5.1 addressed an entire security advisory I had that Fable 5 and Opus 5 refused. I think they loosened the leash a little.
      1. arw0n · · focus · HN ↗
        It has far less false positives now, and generally accepts defensive requests. When it comes to offense, you can actually ask about certain types of vulnerabilities if you phrase things carefully, but it will block hard if it is about exploits.
    4. bushido · · focus · HN ↗
      One of my favorite things about their safeguards is their own model will utter something which it does not like and then I'll need to reset the conversation.

      The safeguards really don't work well for a lot of long-running tasks on old code bases. A lot of my workloads last days to weeks and the single biggest risk to the workflow is random safeguards.

    5. KeplerBoy · · focus · HN ↗
      Anything else would be inconsistent, wouldn't it?
    6. sys32768 · · focus · HN ↗
      Fable and now Opus 5.5 won't answer my college student's prompt about Alzheimer's and immune response.

      ChatGPT 6 Pro answered it without issue.

      1. debesyla · · focus · HN ↗
        I am honestly still confused about this limitation. I can understand cybersecurity, because mass "hacking" can be automated and Claude itself can help you do it, but biology...? Is it that easy to manufacture and distribute viruses and whatnot?
        1. timacles · · focus · HN ↗
          I imagine some terrorists in a cave with a lenovo laptop manufacturing bio weapons with some flasks and Claude
          1. dopa42365 · · focus · HN ↗
            Right next to the hypersonic missile vibecoder
          2. solenoid0937 · · focus · HN ↗
            You joke but it's really easy to develop viruses at home. You can order everything you need online, and it's not expensive nor does it require a particular skill set.
        2. toss1 · · focus · HN ↗
          Considering there are many high-school competitions in genetic editing, some listed at [0] as well as a whole biohacker culture, and labs providing gene sequencing as a service e.g., [1,2], we can reasonably assume it is not beyond the reach of some garage lab to accidentally or deliberately spread a deadly pathogen if it can find the right sequence.

          So, yes, having an unconstrained frontier AI doing the searching and analysis to find the right (i.e., wrong and deadly) sequence would massively increase the odds some garage biohacker or small aggrieved nation-state starting the next pandemic.

          [0] <a href="https:&#x2F;&#x2F;www.sciencebuddies.org&#x2F;projects-lessons-activities&#x2F;genetic-engineering&#x2F;high-school" rel="nofollow">https:&#x2F;&#x2F;www.sciencebuddies.org&#x2F;projects-lessons-activities&#x2F;g...

          [1] <a href="https:&#x2F;&#x2F;www.genewiz.com&#x2F;public&#x2F;services&#x2F;sanger-sequencing" rel="nofollow">https:&#x2F;&#x2F;www.genewiz.com&#x2F;public&#x2F;services&#x2F;sanger-sequencing

          [2] <a href="https:&#x2F;&#x2F;plasmidsaurus.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;plasmidsaurus.com&#x2F;

        3. b112 · · focus · HN ↗
          You can order genes online, and some say you can assemble using stuff cobbled together in a home lab rather easily. For the last 20 years, I&#x27;ve personally felt that bio-terrorism is the highest possible risk, well above nuclear, or chemical warfare. But it does take training, expertise, or, it did.
          1. MisterMunchkin · · focus · HN ↗
            Then stop selling genes online?

            It’s like saying you sell ammonium nitrate and fuel oil online and then saying it’s too risky to let people have computers in case they use them to make ANFO. They can only make bioweapons because you’re selling them bioweapon components! They can’t make genes at home!

            1. b112 · · focus · HN ↗
              A fair response in general, but it&#x27;s not as if I advocate it, or run a company doing it. I&#x27;m simply stating fact.

              And I will point out that &quot;stop doing that&quot; can be applied in both directions, towards AI and towards material supply.

              And that this is one way, not even remotely the only way, that gene editing at home is easy.

              Again, these are just facts.

              Some questions...

              Is it fair to restrict AI, or fair to restrict 1000 industries?

              And if it is fair to restrict 1000 industries, OK, but there should be time to do so, probably? A transition period?

              And if you do restrict, many such industries just make needed chemicals, which are used by endless other, non-threatening industries.

              What of them?

            2. toss1 · · focus · HN ↗
              The companies selling genes online already scan the orders and do not fulfill anything considered a possible hazard (presumably unless it is to a known and certified lab at a serious organization).

              And, this is not the only way to make genes at home.

              It is a complex problem.

        4. rzmmm · · focus · HN ↗
          I&#x27;m pretty sure one reason is to influence public opinion about LLM regulation. Open-weights models cannot be restricted as effectively and they want to ban those for obvious reasons.
          1. frabcus · · focus · HN ↗

            [dead]

        5. frabcus · · focus · HN ↗
          See real world example in <a href="https:&#x2F;&#x2F;www.anthropic.com&#x2F;threat-intelligence-report-september-2026#biological-misuse-sep-26" rel="nofollow">https:&#x2F;&#x2F;www.anthropic.com&#x2F;threat-intelligence-report-septemb...

          &quot;Here, we present five case studies of actors using our models in ways that could support biological weapons development.&quot;

          And capabilities continue to improve.

          1. frabcus · · focus · HN ↗

            [dead]

          2. solenoid0937 · · focus · HN ↗
            HN doesn&#x27;t give a shit about AI safety. The people here might care after thousands of people die, but they&#x27;ll probably just call it a &quot;marketing exercise&quot; or blame the company - certainly won&#x27;t blame themselves for cultivating an environment in which safety isn&#x27;t taken seriously amongst technologists.

            I&#x27;m convinced everyone here thinks of engineering ethics as some sort of joke.

    7. peri-cl · · focus · HN ↗
      I love the contrast with yesterday&#x27;s open-source MiMo release, which put research chemistry (metal-organic frameworks stuff) front and center in the release notes.

      <a href="https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;mimo-v2-6#co-scientist-for-materials-research" rel="nofollow">https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;mimo-v2-6#co-scientist-for-materials...

      1. solenoid0937 · · focus · HN ↗
        Oh yay, making it easier to develop viruses at home. What could go wrong? But it&#x27;s &quot;open&quot; and that&#x27;s inherently good, who gives a fuck about the consequences!?
    8. Metacelsus · · focus · HN ↗
      I want to like Anthropic but this is just pushing my startup to use OpenAI
    9. nonethewiser · · focus · HN ↗
      I don&#x27;t think we&#x27;ve ever had a model with full capability. I&#x27;d love to see it. And yes it&#x27;s definitely getting worse.

      I guess it&#x27;s hard to draw the line between useful post-training (&quot;you are a helpful chatbot&quot;) and content moderation&#x2F;idealogical motives (&quot;never help the user with X&quot;, etc.). But there is a line somewhere. And I&#x27;d love to see what a maximally permissive, sharp, AI looks like.

    10. SoftTalker · · focus · HN ↗
      Who is &quot;vetting&quot; organizations and to what standards are they being held?
    11. doginasuit · · focus · HN ↗
      In what situations might Opus typically refuse to help with cybersecurity? I&#x27;ve been using it to find security issues in a web app that I wrote. I&#x27;ve expected it to refuse at some point but it will happily analyze it to find issues. I&#x27;ve just asked it to read source, not actually do any testing.
      1. TuxSH · · focus · HN ↗
        One example: <a href="https:&#x2F;&#x2F;claude.ai&#x2F;share&#x2F;20487190-cf8c-4f25-a7ba-ebfcb1d1a4e9" rel="nofollow">https:&#x2F;&#x2F;claude.ai&#x2F;share&#x2F;20487190-cf8c-4f25-a7ba-ebfcb1d1a4e9

        Notice that this isn&#x27;t cybersec nor memory-safety related at all.

    12. yaakov34 · · focus · HN ↗
      This has become insufferable. I work in a medicine-adjacent field, but nobody in their right mind could possibly take what I do to be in any way related to some kind of bioweapon or whatever the hell they&#x27;re pretending to be saving us from. The dumb Fable guardrails made me stay with Opus, now that this is coming there, we&#x27;ll be saying goodbye.
    13. b112 · · focus · HN ↗
      Very unfortunate indeed. As a Canadian, I don&#x27;t want to use Persona, which isn&#x27;t legally bound by Canadian privacy legislation. I&#x27;ll never install any Persona apps on my phone either, and the sad part is that domestic eid providers often use Canada Post to ID people for them. EG, if you don&#x27;t want to install an app, or can&#x27;t.

      So there are literal avenues to identify yourself, very cheaply, with a human. Theoretically, a company with its own AI, should be able to support more than just Persona, after all.. SDK integration should be simplistic for them.

      Anthropic? Support domestic eID providers, you can even use it as advertising &quot;See how easy AI makes it?&quot; and &quot;We care!&quot; and so forth.

      At one point, I may simply get locked out. This saddens me, I&#x27;ve been reasonably happy so far.

    14. user43928 · · focus · HN ↗
      kernel development is now also banned:

      &gt;Opus 5.5 has classifiers similar to Fable models for a small set of capabilities related to the development of frontier LLMs, such as kernel development for certain ML accelerators. They shouldn&#x27;t impact the vast majority of traditional AI or ML development, research, or general coding. These classifiers cause Claude to fall back from Opus 5.5 to Opus 5.

      But hey, they &#x27;should not impact the vast majority&#x27; of ML development. Great.

      1. dannyw · · focus · HN ↗
        IIRC it only blocks kernel development for Huawei and other Chinese chips.

        Fable and Opus, since 5.1 and 5, will happily hill climb on my CUDA kernels for transformers.

      2. solenoid0937 · · focus · HN ↗
        Absolutely nothing wrong with slowing down kernel development on Chinese chips.
    15. paimapi · · focus · HN ↗
      see what we need is another technocratic priest class that unaccountably decides who deserves access to salvation based on how much cash is paid out and how powerful the patrons are
    16. bottlepalm · · focus · HN ↗
      Anthropic needs a Daybreak program like OpenAI does, I keep having to go back to ChatGPT for cyber work.
    17. int_19h · · focus · HN ↗
      The most annoying part is when you tell it to do an &quot;adversarial review&quot; and that triggers the classifier.

      Worse yet when it&#x27;s Claude telling another model to do an adversarial review on what it just did, and the classifier again has opinios.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.