> ollama's model index typically only distinguishes variants by their size
Don't use ollama. The entire project is just a series of stupid decisions like this.
How they have any credibility after they told everyone they could run Deepseek R1 on their laptop (by giving the Qwen 7B distill the `deepseek` moniker)...
And last time I have checked, you still can't run rerankers with it, yet you can download the models. See issue 3368, 2years old now.
What's the sota? The classic vLLM/llama.cpp? LM Studio? Unsloth studio any good?