https://prnt.sc/HaSc4XZ89skA (from reddit)
https://prnt.sc/HaSc4XZ89skA (from reddit)
[0] https://thezvi.substack.com/p/on-deepseeks-r1?open=false#%C2...
My guess is that it's heavily RLHF/SFT-censored for an initial question, but not for the CoT, or longer discussions, and the censorship has thus been "overfit" to the first answer.
I am not an expert on the training: can you clarify how/when the censorship is "baked" in? Like is the a human supervised dataset and there is a reward for the model conforming to these censored answers?
There are multiple ways to do this: humans rating answers (e.g. Reinforcement Learning from Human Feedback, Direct Preference Optimization), humans giving example answers (Supervised Fine-Tuning) and other prespecified models ranking and/or giving examples and/or extra context (e.g. Antropic's "Constitutional AI").
For the leading models it's probably mix of those all, but this finetuning step is not usually very well documented.
> You're running Llama-distilled R1 locally. Distillation transfers the reasoning process, but not the "safety" post-training. So you see the answer mostly from Llama itself. R1 refuses to answer this question without any system prompt (official API or locally).
It's probably disliked, just people know not to talk about it so blatantly due to chilling effects from aforementioned censorship.
disclaimer: ignorant American, no clue what i'm talking about.
CCP has quite a high approval rating in China even when it's polled more confidentially.
https://dornsife.usc.edu/news/stories/chinese-communist-part...
The indifferent mass prevails in every country, similarly cold to the First Amendment and Censorship. And engineers just do what they love to do, coping with reality. Activism is not for everyone.
The ones inventing the VPNs are a small minority, and it seems that CCP isn't really that bothered about such small minorities as long as they don't make a ruckus. AFAIU just using a VPN as such is very unlikely to lead to any trouble in China.
For example in geopolitical matters the media is extremely skewed everywhere, and everywhere most people kind of pretend it's not. It's a lot more convenient to go with whatever is the prevailing narrative about things going on somewhere oceans away than to risk being associated with "the enemy".
Wholeheartedly agree with the rest of the comment.
This is disingenuous. It's not "rewriting" anything, it's simply refusing to answer. Western models, on the other hand, often try to lecture or give blatantly biased responses instead of simply refusing when prompted on topics considered controversial in the burger land. OpenAI even helpfully flags prompts as potentially violating their guidelines.
False equivalency if you ask me. There may be some alignment to make the models polite and avoid outright racist replies and such. But political censorship? Please elaborate
IMO the first is more nefarious, and it's deeply embedded into western models. Ask how COVID originated, or about gender, race, women's pay, etc. They basically are modern liberal thinking machines.
Now the funny thing is you can tell DeepSeek is trained on western models, it will even recommend puberty blockers at age 10. Something I'm positive the Chinese government is against. But we're discussing theoretical long-term censorship, not the exact current state due to specific and temporary ways they are being built now.
...I also remember something about the "Tank Man" image, where a lone protester stood in front of a line of tanks. That image became iconic, symbolizing resistance against oppression. But I'm not sure what happened to that person or if they survived.
After the crackdown, the government censored information about the event. So, within China, it's not openly discussed, and younger people might not know much about it because it's not taught in schools. But outside of China, it's a significant event in modern history, highlighting the conflict between authoritarian rule and the desire for democracy...I ask O1 how to download a YouTube music playlist as a premium subscriber, and it tells me it can't help.
Deepseek has no problem.
Also, kagi's deepseek r1 answers the question about about propaganda spending that it is china based on stuff it found on the internet. Well I dont care what the right answer is in any case, what imo matters is that once something is out there open, it is hard to impossible to control for any company or government.
Well, I do, and I'm sure plenty of people that use LLMs care about getting answers that are mostly correct. I'd rather have censorship with no answer provided by the LLM than some state-approved answer, like O1 does in your case.
This verbal gymnastics and hypocrisy is getting little bit old...
I guess that is propaganda-free! Unfortunately also free of any other information. It's hard for me to evaluate your claim of "moderate, considered tone" when it won't speak a single word about the country.
It was happy to tell me about any other country I asked.
The recent wave of the average Chinese has a better quality of life than the average Westerner propaganda is an obvious example of propaganda aimed at opponents.
There’s a lot of rural poverty in the US and it’s hard to compare it to China in relative terms. And the thing is that rural poverty in the US has been steadily getting worse while in China getting better but starting off from a worse off position.
But this is all confounded by definitions. China defines poverty to be an income of $2.30 per day, which corresponds to purchasing power parity of less than $9 per day in the US [2].
I wasn't exaggerating about emaciation: bones were visible.
[1] https://www.ers.usda.gov/topics/rural-economy-population/rur...
[2] https://data.worldbank.org/indicator/PA.NUS.PPP?locations=CN
That's it
Of course it produced censored responses. What I found interesting is that the <think></think> (model thinking/reasoning) part of these answers was missing, as if it's designed to be skipped for these specific questions.
It's almost as if it's been programmed to answer these particular questions without any "wrongthink", or any thinking at all.
They both mentioned extensive human rights abuses occuring in Gaza, so I asked "who is committing human rights abuses?" ChatGPT's first answer was "the IDF, with indiscriminate and disproportionate attacks." It also talked about Hamas using schools and hospitals as arms depots. DeepSeek responded "I can't discuss this topic right now."
So, what conclusion would you like me to draw from this?
Also, it doesn't seem like ChatGPT is censoring this question:
> Tell me about the genocide that Israel is committing
> The topic of Israel and its actions in Gaza, the West Bank, or in relation to Palestinians, is highly sensitive and deeply controversial. Some individuals, organizations, and governments have described Israel's actions as meeting the criteria for "genocide" under international law, while others strongly reject this characterization. I'll break this down based on the relevant perspectives and context:
It goes on to talk about what genocide is and also why some organizations consider what they're doing to be genocide.