logoalt Hacker News

wongarsu • today at 6:11 PM • 2 replies • view on HN

"Sonnet 5.5’s cyber capabilities are a large improvement over Sonnet 5’s, so we’re deploying it with safeguards similar to those on Opus 5.5. Users can still find and fix bugs in their code as part of routine software development, but higher-risk cybersecurity tasks will visibly fall back to Sonnet 5

Sounds like at least for Anthropic models we reached peak cyber capabilities with Opus 4.8. Everything after that falls back to worse models


Replies

ttul • today at 6:13 PM

Daybreak Blue is not bad and the bar to get into OpenAI's program is reasonable.

➕ show 1 reply
gozzoo • today at 7:37 PM

what is the easyest way to use the chinese models and which harness does work with them well?

➕ show 2 replies