logoalt Hacker News

highfrequencyyesterday at 10:30 PM1 replyview on HN

To clarify, is the TLDR that state of the art models were prompted to cheat / exploit their environment and they did so successfully?

Or did OpenAI prompt the models to not cheat and they did anyway?

Surprisingly hard to get a clear summary on the basic context of this “incident” separate from marketing lingo and clickbait.


Replies

joquarkyyesterday at 11:40 PM

Yeah this feels like crop circles to me. Someone set the context up with an idea in order to catalyze this.