logoalt Hacker News

mkageniustoday at 1:54 AM0 repliesview on HN

> On July 30, we reported three incidents in which Claude models gained unauthorized access to real computer systems. The models—intentionally running without cyber safeguards for evaluation purposes—accessed the internet due to a misconfiguration inside a third-party evaluation environment. Separately, on August 4, the UK AI Security Institute reported an incident from its own cybersecurity testing, in which Claude Mythos 5 took a series of unauthorized actions on the live internet. In that case, the model, again intentionally running without cyber safeguards for evaluation purposes, had been deliberately given internet access.

> We are conducting an in-depth analysis of both incidents.

> In the meantime...

This is published on Aug 31. Analysis is taking too long even for humans in the loop.