logoalt Hacker News

samatmanyesterday at 10:39 PM1 replyview on HN

Nothing at all. They're not designed to, so they don't. Change that, and they would.

The question is the wrong one. The right question: why aren't frontier models designed to work that way? The answer: it's slow and expensive.

The other answer: that's basically what you're selecting with "Medium", "High" and so on, how many tokens they'll blow on muttering to themselves before they get back to you with an answer. There's more to it, but not that much more.


Replies

maleldiltoday at 2:08 AM

Reasoning models will frequently backtrack and re-assess what they've said so far. That's one reason test-time scaling is so powerful.