logoalt Hacker News

alkonauttoday at 10:31 AM3 repliesview on HN

There is no relative/percentage increases noted (understandably). Just because i'm lazy: roughly how much more expensive is it to work with v4 flash and v4 pro through the API, compared to before the price increases? Is it 2x, 5x, 10x higher?


Replies

zupa-hutoday at 10:48 AM

# Flash, off-peak

    cache-hit 2.5x
    cache-miss 1.57x
    out 2.36x
# Flash, peak

    cache-hit 5x
    cache-miss 3.14x
    out 4.71x
Edit: fixed the numbers and formatting
embedding-shapetoday at 11:16 AM

Someone made a comparison yesterday, including relative increases, and GPT-5.6 Luna, then later someone also added more OpenAI, Anthropic, K3 and GLM 5.2: https://news.ycombinator.com/item?id=49286679

Already outdated though I think, as GLM 5.3 is latest now :)

floppydtoday at 10:33 AM

About 2x-2.5x off-peak for Flash, 2x-4x I'd say for Pro (x6 on cache in, the biggest increase throughout the board). And twice as much in peak hours.

show 1 reply