so looking at the start docs, you should integrating the same llamacpp setup where they just bake in huggingface support.
second, don't use vague names of models; point to the actual repos you're testing on huggingface. There's enough diversity and specialization that even if you're smart enough to know that colibri is doing something that's particularly applicable to a type of model, it's easy to get lost in all the acronyms.
third, this looks like a fun tool to unite the diversity of random hardware people have, which is always going to win for local inference.