People are desperately trying to cope themselves into thinking that there are alternatives to scaling up transformers to AGI/actual competition to OpenAI or Anthropic. Jev, continual learning, linear attention, local models, non-transformer architectures etc. Imo these are just random technologies that nerdsnipe your average twitter or hackernews user and give them some hope that some underdog can take a slice of the pie.
In reality, none of these really matter. The frontier labs can easily do something like this but likely havent because the size of this market is too small and it is not on the critical path to AGI.
When you have as many resources as OpenAI and Anthropic, theres basically no point in putting compute towards random bets that don't have a predictable return. And at this point, scaling up transformers is almost a surefire way ot putting money in via training and getting money out via increased capabilities AND it speeds up your own business by factors of X. Sidequesting a Jev like product is falling for twitter hype and is likely not going to happen, definitely not by Anthropic, and I'd bet probably not by OpenAI either.