logoalt Hacker News

SwellJoetoday at 5:48 PM0 repliesview on HN

Jebus, that is some sloppy prose. Can people not even be bothered to write the summary themselves, anymore? AI doesn't want anything, so they can never have a point of view, so their prose rambles incoherently across all the various prompts they've seen in a project. This project sounds like the ramblings of a crazy person. Even though the fact that DRAM can do any computation is interesting, nobody should have to read this mess.

"why are we still transferring data to the compute? Why not execute AI inference natively within the memory?"

You already answered that question: 47.5 seconds per token from a tiny 2B 1-bit model model.