Very impressive score for the size, though token use is higher than k3 and far higher than proprietary models, and its price to performance isn't all that far ahead of k3 as a result
>token use is higher than k3 and far higher than proprietary models
GLM sets effort to max by default historically.
>token use is higher than k3 and far higher than proprietary models
GLM sets effort to max by default historically.