logoalt Hacker News

minimaxiryesterday at 7:47 PM6 repliesview on HN

Hy4 apparently has ludicrous traction on OpenRouter already (https://openrouter.ai/tencent/hy4-preview), with trillions of tokens processed in a couple days: more than GLM 5.3 in a week. That said, it's relatively cheap with a 5% cache cost when everyone is still doing 10%/20% cache costs, so Hy4 may be more compelling.


Replies

martinaldyesterday at 8:43 PM

I wrote about this a couple of weeks ago. It's actually often the biggest cost and it tends to be hidden away on most platforms!

https://martinalderson.com/posts/watch-out-for-cache-read-co...

Btw I still haven't came across any decent model that is <$0.01/MTok cache costs apart from deepseek thru their official API (even with the price increases).

Seems like a bit of an opportunity for someone to take - drop cache read costs significantly.

show 1 reply
joegibbstoday at 1:06 AM

If you’re Tencent you can just plug it into some field somewhere that lots of people see right? Like how Meta could put their model on Instagram search

redox99yesterday at 10:57 PM

It's very likely tencent games those stats, buying their own tokens.

Dinuxyesterday at 8:45 PM

Which explains why almost none of my request go though

npnyesterday at 8:29 PM

[flagged]

cyanydeezyesterday at 8:08 PM

i'd be curious if openrouter is just being gamed by these publishers by paying for the exposure.

wouldn't trust they dont do Capitalism like the rest of the AI field.

show 2 replies