logoalt Hacker News

anvuongyesterday at 5:05 PM1 replyview on HN

Not really. The gpt-oss models were notoriously hard to remove the built in safety. It totally depends how it was trained.


Replies

adrian_byesterday at 5:11 PM

Also from the recent Kimi models it seems that it was difficult to remove the censorship.

But in both cases, eventually uncensored variants were published, even if it took more time than for other models.