logoalt Hacker News

manquertoday at 2:56 AM1 replyview on HN

Not necessarily, there could be diminishing returns on mere parameters count .

There is nothing to say for example a 1 Quadrillion parameter model will be vastly more intelligent than current SOTA especially since new training data is largely synthetic today


Replies

HDBaseTtoday at 3:12 AM

That's precisely what he is saying, there is diminishing returns (or optimization left on the table).

show 1 reply