logoalt Hacker News

anon373839today at 5:12 AM1 replyview on HN

> If someone put in frontier AI models from like .... last june I guess? in a box and let me run it with "decent" token throughput I would be happy.

You can have that! Qwen 3.8 Flash-Next is ~Opus 4.6 and runs nicely on a DGX Spark. And that’s just an architecture preview. The Qwen 4 family is expected to arrive this fall.


Replies

rtpgtoday at 5:53 AM

DGX Spark is a biiiiit costly but neat to hear!

Do you know what kinda throughput you’re getting on that kinda setup?

(I have a secondary problem of being “locked into” Claude Code by it being good enough for me, I’d probably need to investigate the other harnesses… my impression is other harnesses are a bit more aggressively OK with nuking your setup from orbit)

show 1 reply