Does this imply that it was a one shot prompt with ChatGPT Pro style models (i.e. best-of-N), rather than the agent swarm approach that was used for Navier-Stokes?
It doesn't imply that, it's just measuring the amount of compute.
It doesn't imply that, it's just measuring the amount of compute.