logoalt Hacker News

znnajdlayesterday at 5:58 PM2 repliesview on HN

Have you actually used the latest Chinese models? Because, in my experience, they're not just cheaper, they're actually better in many cases. In one recent experiment I did a few days ago on a task that I need in production at scale at my company, Qwen 3.8 and GLM 5.3 Flash were not just cheaper, the results were significantly higher quality than GPT 5.6 Sol.

Even if China doesn't continue to release this stuff for free, I think somebody will. Eventually, maybe Europe or maybe a smaller country that picks up this knowledge will. Or maybe just some random philanthropic billionaire.


Replies

switchbakyesterday at 11:24 PM

Yes, I use the messing Chinese models extensively. Like you, I very much appreciate them and sometimes prefer them.

But there is a bar of complexity at which they fail, and repeated invocations typically doesn’t make much progress. This is a small subset of most work, but it still exists. And yes you can help it along, but in those cases I’d typically break out the big guns.

villishyesterday at 7:11 PM

> Qwen 3.8 and GLM 5.3 Flash were not just cheaper, the results were significantly higher quality

How can they be higher quality than models they were distilled from?

show 1 reply