logoalt Hacker News

Pestolast Saturday at 6:59 PM1 replyview on HN

I think the biggest problem with Chinese models is that they seems to overthink for most of the tasks, especially for smaller ones. The OpenAI models have in my experience only gotten better in terms of efficiency.


Replies

fastballlast Saturday at 7:23 PM

Yes, this (imo) is a clear result of benchmaxxing. You can get a much better score on most "intelligence" benchmarks by massively over-saturating reasoning. This looks good on those, but for actual daily usage makes the models much less effective: I don't want a model I use for coding to burn a bunch of reasoning (read: time) on trivial tasks.

show 2 replies