Interesting to see where local models are going to be in the coming days. I am already starting to believe open source models are the way to go in the coming days. With Qwen 3.8 Max, Kimi K3 etx already delivering at part perf with frontier models, the future is going to be exciting.
I purchased a 5070Ti (16GB nVIDIA GPU) a few months ago, and it is absolutely incredible what can be accomplished on local hardware (whether offline or not).
Don't forget `mistral-small` (Apache's LLM), which to me is equally as impressive as qwen3.5 (only benefit of qwen is seeing the pre-text reasoning is often more helpful than the actual text output).