> China may be subsidizing this for now in a way that US companies can't or won't
They're subsidizing this in many ways - Huawei chips, new DDR5 memory fabs, etc.
Ultimately, DeepSeek's architecture is significantly more cost effective than anything from Google, OpenAI, or Anthropic.
Presumably, they'll incorporate DeepSeek's MLA* architecture to get all the benefits for next year's releases (if not this year's upcoming releases) which will bring down their costs...
They need to actually make money, though, so that might still not give them enough room to make enough money.
Ultimately, hardware depreciation is like 80% of total spending. So power is not as big of a deal in cost. The bigger problem is if you can get the power at all, not how expensive it is.
If you want to bring down inference costs, using less hardware is far more effective than getting cheaper electricity.
Google is in a sweet spot, because they aren't paying 80% margins to nVidia for hardware. So they're probably paying half as much deprecation as everyone else is (or maybe 1/4th for inference - which is now the biggest percentage overall).
What’s the TLA architecture? I haven’t read about that.