I've been using the last Deepseek Flash update for a week and I'm amazed. It was a capable model for easy tasks but now it looks like it can do some heavy development for peanuts.
I can't wait to try this new one.
Worse than Luna but more expensive than Luna. Sticking with Luna without sending my data to Deepseek (China)
Currently burning money quickly on official deepseek api. They are also increasing pricing starting today. V4 Flash 0731 still feels like the most outstanding model of the past few months and probably to come.
Still behind Kimi-K3 in almost half of the benchmarks
What I care about is whether the model is capable of the tasks I give it at the lowest cost. Right now I'm using Kimi-K3/GLM-5.2/Minimax. Sonnet is great but I burn through the tokens too fast. Opus 5 set to max is amazing and more intelligent than all of us. .998 of the time I don't need that kind of intelligence. I just need the job done.
https://api-docs.deepseek.com/quick_start/pricing/
Competitive with opus 4.8 but weaker than sol or fable. About 20x cheaper.
@dang - Pls merge this with https://news.ycombinator.com/item?id=49274018
I find it interesting how much adoption seems to be influenced by momentum. Some of these Chinese models are surprisingly capable, but developers often default to the models that are already established as the “industry standard
Benchmarks:
Source: https://reddit.com/r/LocalLLaMA/comments/1vmi0fg/deepseek_v4...