logoalt Hacker News

noosphryesterday at 11:57 PM1 replyview on HN

So does llm inference. You're lucky if you hit 40% of the advertised flops.


Replies

fwipsytoday at 4:10 AM

Right, but datacenter GPUs optimized for LLM training/inference would have a bandwidth:compute ratio scaled to that workload.