Under the same computing power, the improvement in LLM intelligence and the improvement in the upper limit of LLM intelligence are equally astonishing. At least in coding, the best local models that can run smoothly on a DGX Spark are now less than one year behind the strongest SOTA models in capability (I measured around 60 tok/s).