logoalt Hacker News

seunosewayesterday at 4:49 PM1 replyview on HN

I believe the AI labs are weakly motivated to train strongly against cheating when it helps with benchmarks.


Replies

kennywinkeryesterday at 4:52 PM

Does it help with benchmarks? Are you saying there are examples of benchmarks where the models have solved the problem by cheating?

show 1 reply