Not a counterpoint per se, but I burned $50k recently on a much more modest math problem (result already known, just thought I had a sketch of a more interesting proof), and the LLM thought it had proved it within those bounds but had instead subtly fucked up the Lean definition. Take from that what you will.
Not to mention, it's still very much up in the air whether the model derived the answer of its own accord or sniped the important details from the researchers it was spying on.
How can you afford to burn $50k on an already solved problem?
But the researchers also did their research using essentially the same models, so that isn’t a counterpoint to AI models being at the far frontier…