logoalt Hacker News

XCSmeyesterday at 11:12 PM1 replyview on HN

Wait, is it even thinking? Or is it an instant model?


Replies

msdztoday at 3:03 AM

It’s not reasoning, the hardware demo uses a 3.-something generation Llama 8B.

But it’s proven they can automate this (they didn’t etch eight billion weights by hand after all, obviously), so now the interesting question is whether they can scale it to more recent aka bigger models.

After all, there’s already very useful models even for productivity at 27 or 35B.

show 1 reply