logoalt Hacker News

embedding-shapelast Sunday at 11:27 AM1 replyview on HN

I've been playing around with K3 a bunch, but the verbosity of the reasoning makes complete e2e agent work basically cost the same as other smaller models, and I'm not seeing a huge difference in quality, just a way longer e2e completion time.


Replies

sunaookamilast Sunday at 11:44 AM

Same problem with every chinese model currently, they overthink way too much and take too much tokens and time.

show 3 replies