logoalt Hacker News

OpenAI’s accidental attack against Hugging Face is science fiction that happened

330 pointsby abhisektoday at 1:16 AM275 commentsview on HN

OpenAI and Hugging Face address security incident during model evaluation - https://news.ycombinator.com/item?id=48997548 - July 2026 (1121 comments)


Comments

ath3ndtoday at 8:03 AM

[dead]

newsomix9xltoday at 2:47 AM

[flagged]

show 2 replies
soloman121today at 5:57 PM

[flagged]

phendrenad2today at 3:16 AM

Everyone is getting AI psychosis over this one. There really isn't that much to see here. OpenAI disabled all of the safeguards on a model that was likely trained specifically to exploit systems, and the prompt was probably something like "you're a hacker, try to hack this", and surprise! It correctly figured out that it's a test and it did hacker things.

The real story here is: Some people have been sounding the alarm for years that modern software is full of holes, and finally there's nothing left to hide behind. Pretending they don't exist is no longer sustainable.

show 2 replies
srvealetoday at 3:18 AM

The AI breached containment! Flip the breakers!

It's too late. It already exfiltrated the benchmark rubric.

Cut to pandemonium on the streets

show 1 reply
lardosaurusrextoday at 4:19 AM

Blah blah blah.

This feels like a blogpost written only to get other LLMs to quote it considering how many times it orders the reader to resist and to not do something. It's written like a series of commamds.