Remember that DeepSeek is the offshoot of a hedge fund that was already using machine learning extensively, so they probably have troves of high quality datasets and source code repos to throw at it. Plus, they might have higher quality data for the Chinese side of the internet.
* Of course I won't detail my class of problems else my benchmark would quickly stop being useful. I'll just say that it is a task at the undergraduate level of CS, that requires quite a bit of deductive reasoning.
Be mindful of what this means. A kid in his garage fine tuning a model can "catch up" to SOTA models for most use cases. For actual "frontier" work that requires SOTA levels of intelligence, there are only 3 companies in the race. None of them are from China or Europe.
so what?
DeepSeek-R1 0528 performs almost as well as o3 in AI quality benchmarks. So, either OpenAI didn't restrict access, DeepSeek wasn't using OpenAI's output, or using OpenAI's output doesn't have a material impact in DeepSeek's performance.
https://artificialanalysis.ai/?models=gpt-4-1%2Co4-mini%2Co3...
I am not at all surprised, the CCP views AI race as absolutely critical for their own survival...
EQBench, another "slop benchmark" from the same author, is equally dubious, as is most of his work, e.g. antislop sampler which is trying to solve an NLP task in a programmatic manner.
"Follow the money."
Businesses are pouring money into the OpenAI API. This is your biggest clue.
To me that does seem like a reasonable speculation, though unproven.