logoalt Hacker News

behnamohyesterday at 4:49 PM1 replyview on HN

This model doesn’t look solid at all. It comes months after the Qwen model, and in almost half the benchmarks, it performs worse than that. Plus, the next Qwen 3.8 is going to be announced this week. So, this model is DOA.


Replies

dofmyesterday at 5:09 PM

I don't really care that much about benchmarks, but having tested it on one of my puzzle prompts I can tell you that it solves it well, writes clearly, isn't noticeably slower than Qwen 3.6 27B and is much more terse in its reasoning (which will help with preserve-reasoning).

It also has a knowledge cutoff inside this year.

The main limitation is the smaller maximum recommended context.