logoalt Hacker News

GodelNumberingyesterday at 9:09 PM2 repliesview on HN

> That's quite common with many models

Such as?

I can't think of any. Diminishing returns, yes. Occasionally flat, yes. Downright regression, no.


Replies

XCSmeyesterday at 9:15 PM

In my own tests on aibenchy.com, where questions are quite simple, higher reasoning efforts consistently used to do worse than medium for most models.

The reasoning effort should match the complexity of the task against the model's capability.

Hard task with low reasoning = bad

Easy task with very high reasoning = bad

minatoaqua1yesterday at 9:39 PM

grok 4.6