Pretty impressive so far, but needs more testing.
It is more useful than Qwen3.8:27b (which is already quite good) and runs faster on my 7900 XTX / 64 GB DDR4 system.
Local LLM is getting more exciting every day!
Interested to know throughput on 7900 XTX and what setup you're using?
Interested to know throughput on 7900 XTX and what setup you're using?