logoalt Hacker News

walrus01yesterday at 12:12 PM1 replyview on HN

I agree that quantized versions aren't perfect, but using GLM5.2 as an example, the gap between a BF16 and something like a Q8-K-XL as published by unsloth or a similar Q8 quantization is very minimal. For other "large" LLMs there's a fair number of tests showing that Q8 is about 94% as good at literally half the size in GGUF files on disk, and half the RAM usage. Approx. 1500GB for the BF16 vs 820GB for Q8-K-XL.


Replies

aand16yesterday at 5:26 PM

"Very minimal" unless the solution to your current task is in that missing %6 of capability.