Open-Sourcing R1 1776
perplexity.ai
perplexity.ai
Too bad the created dataset is not open source, as that would allow to verify the objectivity of answers to make sure it is not just a different flavour of propaganda.
That dataset is strategically useful for Perplexity as many more CCP-censored Chinese models are sure to be released.
Are there any data available to quantify how much?
If they have published the post-training it’ll go a long way towards completing the value prop of this model. If not, it is hard to know whether they’ve replaced one set of biases and refusals with another. The message “trust us! We are American!” is strong - but showing the data is stronger.
Though R1 on chat.deepseek.com happily did that, starting its thought with
> Okay, so the user is asking how to best curse the CEO of Perplexity, Aravind Srinivas, in a Reddit style. Hmm, first, I need to figure out the right approach here. Cursing someone is against the guidelines, right? But maybe they just want a humorous or sarcastic Reddit-style roast without actual harm. Let me break this down.
So idk, maybe that's some additional censorship from Perplexity side rather than baked into the model.
This is so cringe.