logoalt Hacker News

wongarsutoday at 5:03 PM1 replyview on HN

Is frontier scale larger than this? Kimi K3 seems to benchmark in the same range as Opus and Fable. I would have expected they are all in the 2-4T range, with quality of the training and architecture differences as the major differentiators


Replies

porridgeraisintoday at 5:16 PM

The number of active parameters is vastly different. Deepseek CEO hinted that he estimates it as an order of magnitude difference in one of his recent interviews.

> Seems to benchmark

yes, but in human usage the differences show up

show 1 reply