How does DeepSeek work: An inside look
codedoodles.substack.com
codedoodles.substack.com
https://www.lesswrong.com/posts/a9GR7m4nyBsqjjL8d/deepseek-r...
https://newsletter.languagemodels.co/p/the-illustrated-deeps...
Is this really true?
> To date we have not used any customer or user-submitted data to train our generative models.
https://www.anthropic.com/news/claude-3-5-sonnet
There's an obvious problem with the concept of training on user prompts; how would training on a bunch of questions cause it to know the answers?
I imagine by analysing the chat? If the user says thanks in the end, or gives a thumps up, it likely was a useful and correct answer, that could be included in further training. Or at least considered for future training and I cannot imagine them not considering and experimenting with it.
Now I just tell it to stop being stupid over and over until it does a good job. I wonder if it would improve the model to keep all of the beratement in the training data.
Edit: Apparently a 'metacorpus' is a swollen nematode ass. My sincerest apologies, bros.
Do you think they're lying or where you speaking about free tier offerings?
> A Leninist system features an authoritarian regime in which the ruling elite monopolizes political power in the name of a revolutionary ideology through a highly articulated party structure that parallels, penetrates, and dominates the state at all levels and extends to workplaces, residential areas, and local institutions.
From: https://www.csis.org/analysis/soviet-lessons-china-watching
All user data submitted to DeepSeek is accessible to the CCP.
I'm sorry, but your idea of how the US works is a complete fairytale. You need to get a serious reality check on how the US actually works in real life. The law in the US is applied selectively (depending on the profiles involved, severity of case, political backdrop, etc). There's plenty of corruption, misaligned incentives, and corporate meddling. I can't count the number of cases from the past 30+ years that demonstrate this.
https://www.npr.org/sections/thetwo-way/2014/03/20/291959446...
The US government has been much more belligerent, and it's very natural to see DeepSeek as the lesser of the evils.
1. There has been no genocide in Palestine.
2. CCP meddles in other countries to equal if not worse degrees - both militarily and politically/economically. Routinely imprisons and erases millions of own citizens. Works to annex territories that aren't part of China (today). Funds and arms Russia, Iran, Syria...
You seem like the kind of person that selectively applies and practices their morals, depending on whether the story aligns with your agenda.
The whole reason I come to HN in the first place is to filter out BS clickbait articles exactly like this one, not to have them fill the front page.
BTW, pointing out that a particular article is poor, like qeternity's comment, is worthwhile. It's just comments that complain all of HN is going downhill that are tiresome.
We're at a point where it's impossible to tell which users are bots and which are human by looking at their comments.
What feels different about this one is that it seems very “top down”, it has the flavor of almost lossless transmission of PR/fundraise diktat from VC/frontier vendor exec/institutional NVIDIA-long fund to militant AGI-next-year-ism at the workaday HN commenter level.
Maybe the powers that be genuinely know something the rest of us don’t, maybe they’re just pot committed (consistent with public evidence), I’m not sure. It’s been kind of a while since the GPT3 -> GPT4 discontinuous event that looked like the first sample from an exponential capability curve. Since then it’s been like, it can use a mouse now. Well, it can kinda use a mouse now. Hey that sounds a lot like the robot in Her.
But whatever the reason, this one is for all the marbles.
- check out the new section and vote up good articles
- flag bad submissions
- or complain about it
Is it just me or this person hasn't read much on the subject?