This doesn’t seem obviously true, eg an Anthropic model will never route to Kimi even if it were best suited for a particular task.
Has anyone tried that? I have a feeling that if I put it a prompt Claude would comply. But I am all in on the Claude cool aid.
Why should it? An Anthropic model is architecturally optimized for Anthropic models, routing it to Kimi makes zero sense
[dead]
I think what the parent is saying is that the model itself has the best context for whether a portion of a request should be routed. The specifics of that routing (e.g., should you route to KimiK2) are something that can be trained, finetuned, or even included in a model's startup context.