To clarify, is the TLDR that state of the art models were prompted to cheat / exploit their environment and they did so successfully?
Or did OpenAI prompt the models to not cheat and they did anyway?
Surprisingly hard to get a clear summary on the basic context of this “incident” separate from marketing lingo and clickbait.
Yeah this feels like crop circles to me. Someone set the context up with an idea in order to catalyze this.