logoalt Hacker News

Legend2440yesterday at 7:35 PM3 repliesview on HN

Proof?

In my experience modern models are better at all tasks than models from two years ago, especially complex multi-step tasks.


Replies

nostreboredyesterday at 9:19 PM

For customer support I don't think models have gotten better since gpt-4.1. The class of small models, with limited to no reasoning, that need to handle a complex issue with a touch of empathy, has not improved much.

I think most are actually worth, as agentic harnesses seem to optimize for solving poorly described problems rather than following complex procedures as written. In other words, instruction following maximizing models seem to make worse free-form agents, but they're really all that some domains need.

show 1 reply
alightsoulyesterday at 7:41 PM

you are working on coding. they are working on things like "creative writing" remember that gpt 4o was popular among those who had ai as a romantic partnet?

show 4 replies
criddellyesterday at 8:53 PM

Do any of the big AI companies have a model that are good at tasks that require learning?

For example, every day people teach teenagers how to drive and with only dozens of hours of practice, they are on the road.

show 1 reply