logoalt Hacker News

WarmWashyesterday at 3:47 PM2 repliesview on HN

The benchmark also doesn't include speed. You almost think something has gone wrong when using it because it returns full responses so incredibly fast.


Replies

scrlkyesterday at 3:53 PM

Not just speed, also reliability. IME, Gemini's speed and quality doesn't degrade badly during weekday working hours compared to OAI, and especially Anthropic.

show 1 reply
sotixyesterday at 11:06 PM

This one uses that as a priority weight: https://winstonrc.github.io/ai-coding-agents-leaderboard/