logoalt Hacker News

scrlkyesterday at 10:50 AM10 repliesview on HN

Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion?

EDIT: An open weight version of Muse Spark 1.2 is going to be released as well:

https://x.com/alexandr_wang/status/2086756152034066792

https://xcancel.com/alexandr_wang/status/2086756152034066792


Replies

pu_peyesterday at 11:33 AM

Based on the benchmarks, it seems that Muse Glimmer barely edges out against Qwen3.6 27B, except for tool-calling skills (MCP, etc.). I wouldn't be surprised if they released it now because they are afraid they wouldn't beat Qwen3.8 27B.

show 2 replies
karimfyesterday at 11:32 AM

Yes, and also waiting for the next iteration of Gemma. Muse or Qwen are optimized for coding, while IMO Gemma is still better for non-coding tasks.

https://x.com/osanseviero/status/2086107547535122767

show 2 replies
overfeedyesterday at 7:01 PM

> Will be interesting to see how Qwen3.8 27B compares against this once it releases this week

Considering that Meta distills Qwen[1] (and should!), it'd be hilarious if Muse loses the head-to-head; the "distillation attack!!1!" people claimed distillation on release n-1 is enough to match the intelligence of the latest version.

1. They wrote a paper about it

wronglebowskiyesterday at 10:52 AM

It’s really interesting timing, Qwen over thinking is what kills it for me. I’m just glad we have more options in this size class now.

show 3 replies
Gecko4072yesterday at 11:24 AM

Makes me feel hopeful. Things felt more positive around the llama 3 era. Now it’s like a dark, dreadful race.

imilevyesterday at 12:02 PM

yes i think everyone is waiting to see that ;d, i've been on qwen for the last year and a half now.

spwa4today at 6:43 AM

Well, it has to, since even the MoE models can't really hold a conversation.

aruggirelloyesterday at 2:45 PM

> Seems like dense 30B is back in fashion?

Huh, well... no? Gemma A4B and Qwen A3B are quite popular in fact. I'm sure 3.8 35B A3B will outperform 3.6 27B by all metrics

show 2 replies
ignoramousyesterday at 11:48 AM

> Seems like dense 30B is back in fashion?

Surprising that Meta don't host this model, even as rate-limited free-tier.

> open weight version of Muse Spark 1.2

Wait. Is this "version" different from what Meta serves?

lostmsuyesterday at 11:22 AM

It seems worse than 3.6, but a bit smaller.

UPD. was wrong on smaller, it's actually much larger

show 1 reply