Now that they have reached the frontier in raw performance, I would like to see Chinese models improve their reasoning efficiency.
For all the talk about over reasoning, K3 on low thinking has been rather nice
[flagged]
For all the talk about over reasoning, K3 on low thinking has been rather nice