logoalt Hacker News

Topfiyesterday at 7:20 PM3 repliesview on HN

Unless I have read over it, besides the animation in the intelligence vs speed graph which only mentions internal data and not whether they truly reran the AA suite, there is no actually solid statement on the important aspect of performance.

Neither the Cerebras or OpenAI post [0] outright state that this performs exactly the same as regular 5.6 Sol. I feel if this was 1:1 just Sol but much faster, they'd (rightfully) scream that off the rooftops. A line such as "this is the same performance, just faster, with no downsides" would go a long way in clarity and communication. Along with no pricing information, I'll hold out on further information.

[0] https://openai.com/index/previewing-ultrafast/


Replies

Scaevolusyesterday at 7:33 PM

"delivering up to 750 output tokens per second and without any quality compromise" seems pretty definitive.

show 3 replies
beeringtoday at 12:14 AM

You (or anyone else) can just benchmark and compare. If they were serving a dumber model it would be trivially detectable.

conceptionyesterday at 10:16 PM

This is what Cerebras does- take other people's models and run them very very fast.