logoalt Hacker News

jumploops • today at 8:11 PM • 0 replies • view on HN

If the Terminal Bench 4.0 scores are to be believed[0] GPT-6.1 is an incredibly efficient model.

Yes, benchmarks aren't real work blah blah, but the delta here is so large compared to Astra, it makes it seem like this is distilled Bel or similar.

[0]https://x.com/thsottiaux/status/2105007628460109953