Training data, training methodology. All NOT OPEN.
Until we know what a model is trained on, and how it is trained in high detail, I hesitate to call them "Open Source" in any way. They are free. But, we don't know what their priorities are etc. Witness the censorship we see in all models in one form or another. I'm not absolving any side of this.
Just saying: Don't be blind.
Did you even read my comment? They explicitly DO share their training methodology in depth in Technical Reports on arXiv.
DeepSeek completely revolutionized LLMs and every western LLM today uses or is inspired by the their innovations including Group Relative Policy Optimization and Multi-head Latent Attention.