logoalt Hacker News

ljosifovyesterday at 8:33 PM1 replyview on HN

Hear hear. IQ tokens to cheap to meter upon us. So many things changed since last week. Now I've had Prime agent session grinding into its 20-th hour still not giving up. Been using opencode-go since Go sub appeared. What made a difference was deepseek-v4-flash and mimo-v2.5 showing. Very similar middling models ~300b so light on the gpu. 1M context and hybrid archs - so one can actually make use of that 1M (don't grind to a halt like others). In OMP I have one the primary (default), the other one as /advisor looking over the shoulder and nagging. On opencode-go in credits counting they are the bottom-2 in cost, cheaper by 200-350 times than than the top-1. Last week with deepseek-v4-flash-0731 another jump - now it's closer to the top models then to the middle. Now I don't even need the /advisor probably. Still left it there it's sometime amusing the models back and forth. :-) DeepSeek offer /v1/responses api now with flash-0731, so setup Codex to use that too. I'm loving this :-)


Replies

klardotshtoday at 2:53 AM

I would not recommend DSV4F (even 0731 edition) without an advisor. On its own it’s an absolute drunk intern in my experience, but with an advisor model watching like a hawk when it gets stuck in loops or goes down boneheaded rabbit holes, it’s fine (and very cheap). I’ve been using GLM-5.2 as my /advisor but might try just a second DSV4F instance.