logoalt Hacker News

piyhyesterday at 6:58 PM2 repliesview on HN

Sol on Cerebras is going to be expensive AF


Replies

jaggederestyesterday at 9:42 PM

Is it? I think waferscale might actually be cheaper per-token, it's just so many more tokens, and of course right now it's not a full buildout so the availability is limited as well. I'd imagine they'll be migrating to whichever inference method is least expensive, and I expect asics to be the ultimate answer.

sigmoid10yesterday at 7:32 PM

Moving from either frontier intelligence or frontier latency to a single model that does both at the same time is potentially a game changer in certain industries. I can easily see e.g. hedge funds dropping tons of money on this, because it means they can now do the same thing as their competitors, but much faster. That's basically a license to print money.

show 2 replies