I see 6 alternative providers listed on Openrouter for DeepSeek V4 Pro for example.
I’d rather use the phone home version (deepseeks own endpoint). The benefit is that I’m fairly certain that they actually host the model I’m paying for.
I’m not trying to be negative here, but your point is invalidated by that particular event in itself.
If you have no problems shitting on tens of thousands of authors of books, you don't have problems shitting on your customers as well (which they have proved again and again, see https://en.wikipedia.org/wiki/Facebook–Cambridge_Analytica_d...)
Let's just say I wholeheartedly disagree with your viewpoint and leave it there :)
What I’m trying to say is that EVERYONE uses your data, even the sensitive type. So you might aswell use an endpoint that does what it says and treat EVERY endpoint whether that’s OpenAI or anthropic as if it’s collecting all of your data.
No but seriously, I am astonished by the level of trust you have for these for-profit companies. I’ll remind you of this quote:
”Zuckerberg: People just submitted it. Zuckerberg: I don't know why. Zuckerberg: They "trust me" Zuckerberg: Dumb fucks”
You let us know what your real complaint is about and let's not feign indignation at open models and research.
they took your rights and your data.
Chinese labs took your data, trained their model, and tell you "this paper details how our models are trained using your data, here is the final weights of our model trained from your data, feel free to use it for what you want, it is your model trained on your data".
they converted your data, everything is still in your hand under your control.
you couldn't see the difference?
Your specific question can actually be translated as -
1. why people don't stop Chinese labs so US monopoly can be maintained?
2. why people don't stop Chinese labs providing free models to those who would otherwise never be able to afford the same $200 USD/month Anthropic and OpenAI subscriptions.
3. why people don't complain Chinese labs publishing those trillion dollar secret ideas on model training.
well, because most people are not dickhead I guess?
Right now they are doing that because they are still trying to catch up to Anthropic, Google, and OpenAI.
The moment they have the special sauce, they will shut it down and you won't be able to run their stuff anymore outside of them. Why do I say that? We already have the evidence in the diffusion model arena. All the chinese labs were pumping out open weights models for image and video, the moment they got to SOTA, they stopped doing it. Less and less is being released.
Chinese companies aren't doing open weights models out of the goodness of their hearts, they are doing it because it help their entire industry catch up. Don't get it twisted, this is very much a US vs China battle here. China wants to win and I am not sure how they won't. Deepseek is the first major large model trained on Huawei chips. It won't be the last and I am betting that China will make up for lesser performance of those chips with more manufacturing and power generation.
I am very bullish on China winning the AI war here. But I also am not naive enough to think that the Chinese companies is doing open weights out of wanting to make the world a better place or the goodness of their hearts. It undercuts the american AI companies.
User publishes to github => Deepseek trains with GitHub data => Deepseek gives model away for free => User did not work for Deepseek (in the sense of giving it's labour for Deepseek to make money)
You can use zero data retention and zero training providers for most open weights. See OpenRouter and OpenCode Go/Zen for examples.
This is actually one of the big selling points behind open weights - neither China nor the US get your data.
Seems ok for MIT like licensed code though
We're on the verge of a golden age of software as soon as someone finds a court with courage.
The point is not that this situation seems absurd. The point is that we need some point where we say whats ok or not.
And by ignoring licensing of public code already we moved it closer to the worse end of the spectrum
But a court may differ in the future.
This cute policy of mine won't affect anything though. The more we use the models, the more the models will replace this kind of work. Centralisation of power is inevitable; in Medival Europe, we used to have state & church ruling. In modern times but before the internet, it was probably state and banks. Maybe with ongoing digitization (bank offices disappearing) making banks less costly to operate; combined with with bank bailouts, maybe govenments will fully nationalize or at least banks will consolidate.
Then the AI companies will consolidate with the internet information and communication companies (Google/Meta for the US, and Alibaba/Tencent for China). Maybe we'll end up with a few de-facto governmental megacorps that rule in tandem and close cooperation with the formal government, who might handle mostly infra, utilities and the army. The megacorp would control narrative more and take more of a paternal role (educating and protecting the citizens, normally handled by formal governments).
Does this make sense?
And unfortunately AWS doesn't have prepaid billing, so you can't just give the internet access to your API key without getting FinDDoS'd.
There's some use cases I won't use a hosted model for, and will only do self hosted.
Otherwise, if they're going to keep releasing open-weight models, I'm going to keep giving them data.
Do you really think OpenAI, Anthropic or any other entity in the same business respects your data?
The Chinese AI companies who release open weights actually deserve whatever input you give them. They are the reason why there is competition and not duopolies in the domain.
OpenAI, I wouldn't be surprised if you were right.
But the more important one is the social contract. Github came far before LLM era. The branding around it is being the storage of open source projects and many users want to it stay away from AI hype. You won't expect LLM providers to stay away from AI hype (duh) so it's less an issue for them.
US has too much influence atm. I'm ok with switching between "bullies".