logoalt Hacker News

_ache_today at 2:57 AM1 replyview on HN

If the performances are comparable, and there is no evidence it's not.

in/out ($) Gemini : 1.5 / 9.0 | Qwen 3.8: 0.15 / 0.47

That is a massive cost reduction.

Refs: https://www.alibabacloud.com/help/en/model-studio/model-pric... https://runware.ai/gemini-omni


Replies

killingtime74today at 8:47 AM

You can't just look at the per token cost, but how many tokens it takes on average to do a task. The difference can be massive.

show 2 replies