I've been playing around with K3 a bunch, but the verbosity of the reasoning makes complete e2e agent work basically cost the same as other smaller models, and I'm not seeing a huge difference in quality, just a way longer e2e completion time.
Same problem with every chinese model currently, they overthink way too much and take too much tokens and time.
Same problem with every chinese model currently, they overthink way too much and take too much tokens and time.