NVidia has reasonable incentives to keep things open indefinitely, it wants people to use it's GPUs
Issue is that llama.cpp is the best way to run models on hardware that isn't nvidias.
It'd be really nice if I could use my egpu 4090 on my macbook pro..
Issue is that llama.cpp is the best way to run models on hardware that isn't nvidias.