Henry Yuen's (whose work problem 6 builds on) comments on this are worth reading IMO: https://bsky.app/profile/henryyuen.bsky.social/post/3ms2jpch...
It sounds like he hasn't verified the results of a problem that he has personally worked on, so how many of these problems have actually been verified?
This starts to feel like chess engines. It’s obvious their play is superior but it’s impossible for humans to understand the moves.