logoalt Hacker News

pyraletoday at 5:59 AM1 replyview on HN

There is no way you would run a dense 27b model on that spec. I ran 3.6 27b on a 64gb ram, 24 gb vram, and it felt like the lower limit for this model with a decent context window.

If you want a better experience, maybe wait for either a moe model (like 3.6 35b A3) or a model with less parameters (like 9b). Qwen has been releasing those in the past, so maybe we’ll have them for 3.8 too.


Replies

DanielHBtoday at 9:43 AM

From my experience if it doesn't fit on vram it is rarely worth to bother except for a few narrow tasks.

For example make an essay about something where you don't actively engage with the LLM after the initial prompt. So mostly one-shot prompts.