logoalt Hacker News

dist-epoch • today at 7:48 AM • 2 replies • view on HN

Jev is rumored to be a 30B model, and it's input price is MUCH cheaper than similarly sized models. The maker is also heavily focused on having a profitable product, so it's unlikely to be subsidizing the cost, especially since they say they have more demand than what they can serve.


Replies

jampekka • today at 9:05 AM

You can get Gemma 4 26B A4B at the exact same token input price of $0.042/M. GPT-5 nano is not much more expensive at $0.05.

https://openrouter.ai/google/gemma-4-26b-a4b-it

➕ show 2 replies
imtringued • today at 10:23 AM

They can't overturn the economics of attention by restricting themselves to a single token output.

Sure they are no longer memory bandwidth bound thanks to that but someone could add a similar projector to a conventional model, train with a Jev style dataset and call it a day.

Whatever they are doing on inputs must either mean they intentionally chose a Mamba successor or they suffer from the same compute costs as everyone else.

➕ show 1 reply