Is there anything -- any possible scrap of evidence whatsoever -- that would convince you that this is not merely a marketing scheme?
This is becoming an idée fixe among the HN crowd. Seemingly nothing can dislodge it, no matter how alarming the incident.
GPT-6 could grab the nuclear launch codes tomorrow and there would be a top-voted comment chuckling that it's all some scheme to pump up the IPO.
---
Put another way, how would you have done the write-up about one of these breakout incidents, if you were in an Anthropic/OpenAI employee's shoes, and (by hypothesis) your intent were not "marketing"? And in a way that doesn't trigger the "it's all marketing" HN top-ranking comment?
I would dedicate a portion of my organization to making O.S. tools that protect against and contain AI models
I always find it bizarre how rational thought goes out the window whenever AI is involved in HN. There's gotta be something in the water...
This is published on a marketing website.
If it were not a marketing scheme, they would responsibly disclose the vulnerabilities to the code owners, and go on with their lives.
Here is one piece of evidence that would convince me: they admit they can't contain it, the they erase the weights and dismantle the company.