(comment got too long, continuation)
> All sufficiently capable models, open and closed, should go through mandatory safety testing
Why? And how do you design those tests?
For example, bombing girls school in Iran - is this allowed use according to you or not?
If not allowed use, then how do you guarantee that you don't have a separate agreement with DoW which makes it allowed use and only your model passes it?
Bombing that building was an accident due to it being on what was part of a military target and I'm not sure if it was ever made clear whether the building was technically dual use despite also being used as a school. Either way, bombing a bunch of school girls wasn't intentional and nobody reasonable would assume it was.
There was a time we bombed a bus or van with kids in it, but we admitted it and apologized for the mistake. Nobody wants to be bombing kids, first because they're innocent, but second because there is no military advantage to it since it's bad PR.
Many of these targets were identified before Anthropic or OpenAI even existed.
This isn't Twitter. You can write long comments here.