logoalt Hacker News

bigbuppoyesterday at 3:51 PM1 replyview on HN

Would there be any motivation for the humans behind the scenes to be directing tasks in a certain way knowing that trillions of dollars are on the line? Is it in any particular company's best interest, one that just announced their latest model is "really AGI", for them to be known to have an AI that's just out there trying to escape its confines?

Cui bono?


Replies

bulderyesterday at 9:10 PM

I personally doubt they gave the models specific instructions calling out named websites to communicate over, but I do agree that OpenAI is likely training their cybersecurity-enabled models in ways that encourages abusive and amoral behavior. Either through negligence or by finding it gives them better results.