logoalt Hacker News

Lwrlesstoday at 5:15 AM0 repliesview on HN

I found some on Reddit, and here are the two benchmark results from review sites: - against a desktop 5090: https://nanoreview.net/en/gpu-compare/geforce-rtx-5090-vs-ap... - against a laptop 5090: https://nanoreview.net/en/gpu-compare/geforce-rtx-5090-mobil...

With larger (also really fast) unified memory on the chip, it could easily load larger weights, and the thing is that the quantized results on the M5s are really good, I believe that for most local inferencing, users would be using INT4 (maybe more bits per weight sometimes) quantized models, that might be where the Neural Accelerators kick in. In raw power, a desktop 5090 easily outperforms the M5 Max, but for this specific use case, I believe that the M5 Max is good enough.