It also takes some load off the AI data centers.
IDK if that might be a concern for Apple or their AI partners.
It worsens the supply crunch, no? A unit you use sparingly vs that memory going into a GPU that serves many more people.
Surely this is not something that motivates the vast majority of people using local LLMs.
It worsens the supply crunch, no? A unit you use sparingly vs that memory going into a GPU that serves many more people.