logoalt Hacker News

Schlagbohreryesterday at 8:53 PM1 replyview on HN

Meanwhile those of us with 128GB RAM plus some VRAM don't have any good modern (last 8 months) open weights models to make use of all that. I don't care if it would run 5 tok/s, I want a smarter model than Qwen3.6 which avoids loops and can handle more context than 80k before crashing.


Replies

heysagniktoday at 1:17 AM

why don't you use the quantized version of kimi-k3