For certain tasks, a model like Jev may be intrinsically more efficient than an LLM because it doesn't have to predict a token distribution and can instead focus solely on the probability of a single question/action