logoalt Hacker News

jsnellyesterday at 5:24 PM1 replyview on HN

You should probably look at the cost/score graph by effort level instead:

https://artificialanalysis.ai/models/claude-opus-5-5#intelli...

It is most of the pareto frontier.


Replies

drbsclyesterday at 5:29 PM

Not disputing the increase in quality, just stating that non-cherry-picked benchmarks show it is more verbose at Max effort

show 2 replies