logoalt Hacker News

madducitoday at 5:11 AM3 repliesview on HN

Till now I was using successfully Qwen 3.5 and Gemma 4 at a reasonable speed


Replies

kzrdudetoday at 10:47 AM

I think (maybe I missed something) that identical size and quant versions of Qwen 3.5 and 3.8 should run at the same speed. It’s the exact same architecture.

show 1 reply
pyraletoday at 5:59 AM

There is no way you would run a dense 27b model on that spec. I ran 3.6 27b on a 64gb ram, 24 gb vram, and it felt like the lower limit for this model with a decent context window.

If you want a better experience, maybe wait for either a moe model (like 3.6 35b A3) or a model with less parameters (like 9b). Qwen has been releasing those in the past, so maybe we’ll have them for 3.8 too.

show 1 reply
mobelkhtoday at 5:16 AM

were you running the MoE models? those perform better speed wise