logoalt Hacker News

fermuchtoday at 2:08 AM1 replyview on HN

xhigh tells it to overthink and re check everything. Low tells it to only do the minimum thinking necessary. I would suggest to give qwen medium which doesn't inject any thinking directives into it and also to give as much context as you can, ideally around 500k tokens or even 1M if you can. Big complex tasks like these make the model hit the compaction trigger a lot and they end up re thinking the same thing several times in my experience.


Replies

kennywinkertoday at 3:56 AM

Doesn’t it max out its context at like 256k?

show 1 reply