Hence why we talk about alignment and things like reward hacking. There are lots of people that are saying "if we just .... " the model will be aligned, or that we don't need alignment at all. These people are foolish.