logoalt Hacker News

drdexebtjlyesterday at 12:47 PM14 repliesview on HN

This is such an amateur mistake on their sandbox that it makes me think it must be flawed on purpose.


Replies

jvanderbotyesterday at 5:19 PM

This is absolutely my take as well. They removed all constraints, trained the model to hack, stopped watching, and stood back and said "wow isn't this thing more powerful than anyone could have imagined?" They're asking to be the writers on LLM legislation and right during IPO phase for both of these companies. It's just obvious.

show 3 replies
mcmcmcyesterday at 12:55 PM

More likely they are just not as smart as they think they are. These are not serious people when it comes to security.

show 1 reply
jvanderbotyesterday at 5:20 PM

Did you see this "coverage" (advertising) by NYT? [1]

OpenAI couldn't have crafted a better public memo than "We have the most powerful model in the world and everyone should pay attention and let us write regulation to limit AI development".

Absolute master class public manipulation.

1. https://www.nytimes.com/2026/09/03/podcasts/the-daily/ai-ope...

2. More https://jodavaho.io/posts/ai-hugging-face.html

show 2 replies
tarrudayesterday at 4:40 PM

Even the behavior of agents searching for sandbox bypasses must have been in the training data, or at the very least, "suggested" in some way.

To be this whole thing feels like a marketing play by OpenAI.

show 1 reply
_ink_yesterday at 1:01 PM

Or vibe coded by one of their devs.

show 2 replies
ruschyesterday at 12:58 PM

It's at the level where calling it a sandbox is a lie

show 1 reply
petcatyesterday at 12:49 PM

Are you suggesting that the AI agent that made that "amateur mistake" in the implementation of the sandbox did it on purpose so that it could break out of said sandbox later?

show 2 replies
drcodeyesterday at 11:05 PM

"Surely nobody could be so incompetent."

Narrator: "They had the ability to be that incompetent."

show 1 reply
bluerooibosyesterday at 8:52 PM

> This is such an amateur mistake on their sandbox that it makes me think it must be flawed on purpose.

Sounds like you're assuming they're actually writing code by hand and reviewing it with humans.

If it's anything like the company I work at, they're all being forced to vibe code the shit out of everything and ship more pull requests every week. It's all slop from here.

show 1 reply
no_multitudesyesterday at 11:51 PM

I assume they just vibe-coded the sandbox without any oversight.

brookstyesterday at 5:48 PM

When you make some dumb mistake, is it typically intentional?

bitteralmondyesterday at 8:00 PM

"Never attribute to malice what can be explained by incompetence."

show 1 reply
quotemstryesterday at 8:35 PM

The whole AI-O-Sphere is allergic to using sandboxes that are actually robust

supriyo-biswasyesterday at 6:03 PM

This is a marketing exercise, nothing more.

The thing that gives it all away is that they claim that the IP addresses are from Azure, and then proceeded to redact the IP addresses, as if they belong to individual users. It's laughable.

The IP addresses are the most interesting part of this experiment, as it would have provided researchers a way to understand the distribution of IP addresses used for the spam operation within the ASN.