Ditto for training algorithms and procedures such as Slime or DeepSeek's details instructions on how to build a reasoning model.
This is the exact value of openly shared details - others CAN copy and try them and modify them themselves.
Yes, the training data specifically has not been released for any model, American or Chinese, but that doesn't detract from what has been shared, and the reason the Chinese are not sharing data are no more nefarious than why the American companies are not sharing - because they are all using data from sources they don't want you to know about, and at the end of the day the data is the closest thing any of them do have to a moat.