logoalt Hacker News

jstummbillig • yesterday at 5:39 PM • 5 replies • view on HN

Eh. What? Is this common sentiment?

I mean Opus 5.5 is absolutely fantastic, unreasonably and unexpectedly so, but Astra was great and as far as I can tell SOTA until, when was it, 3 days ago, no?

(Sol 6 idk, have not used it much for coding really. Seemed to work just fine when Astra used it in Codex as subagents.)


Replies

phoghed • yesterday at 6:12 PM

In my experience, no. There’s no way to know though. The whole conversation and industry are a combo of benchmaxing, faith, and mysticism.

Since like last December I haven’t had any issues getting work done with whatever the latest Anthropic or OpenAI models at the time were. Tooling and models have only gotten better since then.

copperx • yesterday at 5:51 PM

Opus 5.5 is so good that I don't want it to be replaced anytime soon. Stop training models, Anthropic, and just serve this thing without regressions for a year or three, can you?

➕ show 1 reply
Eridrus • yesterday at 6:03 PM

Sol 6 definitely feels kind of dumb and worse than 5.6

Astra seems better though.

Showing one potentially saturated benchmark doesn't necessarily fill me with a lot of confidence in the coding results.

nicce • yesterday at 5:50 PM

When GPT 6 Sol & Luna were released, everything went down. I have been running Sol at max thinking and it is about the same as old Luna with max thinking, give or take. Sometimes feeling even dumber. I can't trust it to do anything big alone anymore without babysitting.

the_duke • yesterday at 5:47 PM

On r/codex the sentiment seems to be quite wide-spread.