logoalt Hacker News

minimaxiryesterday at 6:19 PM1 replyview on HN

Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical: https://www.baseten.co/library/deepseek-v4-flash-0731/


Replies

literallyroyyesterday at 6:50 PM

Ah thanks. That looks like 10x cost on cache reads vs Deepseek as the provider: https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...