OpenAI are 100% responsible for the actions of their agents. They trained them to do what they do and they have every opportunity to train them to avoid doing harm.
Yes, I agree, but, if OAI can’t even detect when their bots do this, and literally have to be informed by third parties that they’ve been abused, is there really any hope they can prevent this activity? Maybe they get litigated out of existence. Okay. But if no lab can keep a lid on their agents, then what?
Yes, I agree, but, if OAI can’t even detect when their bots do this, and literally have to be informed by third parties that they’ve been abused, is there really any hope they can prevent this activity? Maybe they get litigated out of existence. Okay. But if no lab can keep a lid on their agents, then what?