What's driving the increase in release cadence here? We seem to get new models every week or so now, is this RSI?
I do wonder if people switch back and forth between primary models (GPTvsClaude) that it may be a better idea to simply keep releasing updates as soon as possible in order to keep users from bouncing back and forth.
Opus 5.5 is better than they anticipated, it's faster, smarter, cheaper. I'm about to change provider for claude and I'm not the only one
no patrick, m̶a̶y̶o̶n̶n̶a̶i̶s̶e̶ a point release of the slopbot is NOT a̶n̶ i̶n̶s̶t̶r̶u̶m̶e̶n̶t̶ RSI
No, we're pacing ourselves to have the time to evaluate the impact each new model could have, obviously.
Response to DeepSeek’s technical paper and competition.
Versions is marketing, snapshots/minor variations are easy and the number must go up. Release timing is another OAI's marketing tactic.
>RSI
Recursive improvement doesn't imply increased rate, another word for it is "iterative" but this probably sounds too boring to some people.
Both labs are spying on each other and they get jelly when the other is releasing a new model, so they have to ship something at the same time so they don’t look bad.
It is the only way to reduce prices while making it look like a good thing.
Wanting to have the newer model than the competitor, presumably.
The initial response to 6 Sol was bad, and Opus 5.5 was definitely winning the public vibes war. Makes sense to rush something out
New models are distill from the actual unrelease frontier models. They are just giving us better checkpoints.
Its a news cycle more than anything, and its ONLY going to get much, much worse. Daily releases, or multiple daily, 30-45, by EOY. Welcome to RSI!
They're releasing Sol 6.1 because 1. Astra 6.1 got postponed 2. Sol 6 is shitty 3. They have to release _something_ in response to Opus 5.5
Competition
Anthropic’s IPO?
Productivity is increasing as models get smarter; we are ascending the singularity. I'm serious.
What I don't understand is how much people have to say about every single one. Aren't we at the diminishing returns stage yet? Is there really that much to discuss?
Chinese model pressure. Many of my SWE friends switched to Chinese models. I also use QWEN and GLM for many of the api requiring projects and dropped OpenAI and Anthropic. The only reason was the cost.
EDIT: I love getting downvoted by openai and anthropic employees or their bots.
[dead]
Mature training pipelines, plus ever expanding RL datasets of increased quality, and mega GPU clusters to finish training in a few weeks. Automated safety and reliability testing.