"Sonnet 5.5’s cyber capabilities are a large improvement over Sonnet 5’s, so we’re deploying it with safeguards similar to those on Opus 5.5. Users can still find and fix bugs in their code as part of routine software development, but higher-risk cybersecurity tasks will visibly fall back to Sonnet 5
Sounds like at least for Anthropic models we reached peak cyber capabilities with Opus 4.8. Everything after that falls back to worse models
Between OpenCode and Openrouter which one would you suggest? Sometimes I have pure grunt work to be done on non-sensitive data for which I want to use Chinese models. For example, tasks like extracting something from publicly available large pdf files.
opencode with model inference on cheaperinference.com has been working well for me - glm-5.3-flash is shockingly cheap (i've spent a total of a few dollars over several weeks of heavy usage), fast and capable for cyber tasks
wongarsu · · focus · HN ↗
Sounds like at least for Anthropic models we reached peak cyber capabilities with Opus 4.8. Everything after that falls back to worse models
gozzoo · · focus · HN ↗
Iolaum · · focus · HN ↗
malshe · · focus · HN ↗
Flere-Imsaho · · focus · HN ↗
malshe · · focus · HN ↗
beveradb · · focus · HN ↗
asp_hornet · · focus · HN ↗
I use GLM directly from z.ai, they do not retain or train on your data accordingly to their TOS.