Meta is rocking AI. As of last week I have been using their excellent muse coding harness with their model Muse Spark 1.2.
Starting this morning I am running their new local 30B model muse-glimmer on my old MacMini 32G using Ollama (remember to increase the context size!) and pi coding harness. I am getting good results with muse-glimmer running locally, with the caveat that everything runs slowly (e.g., give it a task and then go walk outside or do Qi Gong exercises for a while).
seems to underperform on Terminal Bench compared with qwen3.6-27b: 51.7 vs 60.7
Newb question but I’m curious what would help it to run faster? Would it need more vRAM or just system memory?
Friends Don't Let Friends Use Ollama https://news.ycombinator.com/item?id=47788385