I purchased a 5070Ti (16GB nVIDIA GPU) a few months ago, and it is absolutely incredible what can be accomplished on local hardware (whether offline or not).
Don't forget `mistral-small` (Apache's LLM), which to me is equally as impressive as qwen3.5 (only benefit of qwen is seeing the pre-text reasoning is often more helpful than the actual text output).