39 karma · joined July 12, 2016
Please reach out to us if you'd like to look at the dataset.
Happy to send you the dataset if you'd like! Please reach out to our email linked in the post.
If you ran it on your computer, then it wasn't R1. It's a very common misconception. What you ran was actually either a Qwen or LLaMA model fine-tuned to behave more like R1. We have a more detailed explanation in our analysis.
Censoring and straight up propaganda is built into V3 and R1, even the open source version's weights.
We ran the full 671 billion parameter models on GPU servers and asked them a series of questions. Comparing the outputs from DeepSeek-V3 and DeepSeek-R1, we have conclusive evidence that Chinese Communist Party (CCP) propaganda is baked into both the base model’s training data and the reinforcement learning process that produced R1.
We're in the middle of conducting research on this using the fully self-hosted open source version of R1 and will release the findings in the next day or so. That should clear up a lot of speculation.
Meanwhile, we've released the first part of our research including the dataset: https://news.ycombinator.com/item?id=42879698
We evaluated DeepSeek R1 and confirmed that its guardrails deviate significantly from other model providers. We’re currently updating it to behave more in line with Anthropic and OpenAI’s models.