logoalt Hacker News

rustyhancockyesterday at 7:13 AM4 repliesview on HN

He's also completely missed what many of us see as the primary the issue, HF was attacked by a US frontier model.

HF could not be helped by US frontier models because of the "safety" features they have.

HF had to use an open model from china.

Anthropic wants to add those "safety" features to open models - especially from china.

End result would be HF hack would have continued atleast until the Monday that OpenAI engineers finally walked back into work.


Replies

ppap3yesterday at 12:20 PM

I'm very confident that it was staged and coordinated. This propaganda started because they cannot evolve their models further. See Fable and Sol, they are lame. They seem incredible at first but the more you use you can see the trickery. It is a matter of time for someone to prove they are marginally better only because they inject more information to the harness at server side.

show 2 replies
rob74yesterday at 8:09 AM

I case you're wondering, HF = Hugging Face

show 1 reply
imrozimyesterday at 3:43 PM

Worth checking, elsewhere in this thread someone points out glm only assessed the damage after the fact, it didn't stop the attack, and Hf apprently never sought access to a trusted defender program with a closed model either the open model saved them farming might not hold up.

show 1 reply
b112yesterday at 8:20 AM

I think the fair nuance here is, an administration which used Executive Orders to force guardrails, meaning US companies must retain them, even if they might want to drop them now.

And on top of that, with low/no guardrails, people call you a child pornographer(grok), so the public is also against it. Yet mysteriously few complain about Chinese open models being child pornographers.

So even if your goal isn't ethical, but just fiscal, it's reasonable to say there are two standards. And to complaint in some way.

I don't think banning is going to work, that's just silly. And over the next few years, everyone and their dog will have local GPU compute to train locally. People have home labs, the bar isn't that high, and eventually large text datasets will escape from Anthropic and other companies, allowing for comparable training.

It's a genie that's not going back in the bottle, the bottle is smashed.

The only reasonable outcome would be section 230 style carveouts so that there is zero liability for anything a model does.

Because having guardrails on corporate models barely months ahead of open ones, which will never be restricted, is entirely pointless.

show 4 replies