Countries should want control over _where_ the compute is happening rather than _what code_ is running.
What's wrong with a country hosting a Kimi, Qwen or GPT-Oss on their hardware for their government work purpose?
Countries should want control over _where_ the compute is happening rather than _what code_ is running.
What's wrong with a country hosting a Kimi, Qwen or GPT-Oss on their hardware for their government work purpose?
They are not neutral technology, they are a direct representation of the training set that has been chosen and how they are aligned.
In many ways, they are ideology made code.
If we leave building them to the US and China, only their way of seeing things will be digitized.
I don't like the idea of that.
The censorship works kind of like with Fabel, it kicks in before the model responds.
Furthermore, the expertise in designing and training these models is valuable as well. The existing models are good as a starting point in terms of learning from previous mistakes, but we should not just let a handful of American and Chinese people keep the knowledge and expertise.
One problem with this particular project, though, is that copyright has been enforced for Dutch LLM training before, and the AI industry cannot exist without massive scale piracy, the likes of which has never been seen before. A lot of Dutch training material exists in pirated books that AI companies in countries that do not care about copyright have access to, but are exempted from the training set here. The impact of enforcing copyright on an AI model will be quite interesting to see.
As for accessing pii, I imagine the value here is in the fact they're local, which has nothing to do with the "sovereignty" of these models. If anything, a model is more likely to be tricked by a malicious prompt the farther it is from the sota.
Lots of bias towards English sentence structure, idioms, etiquette, etc.
"PewDiePie has built a custom web UI for self-hosting AI models called "ChatOS" that runs on his custom PC with 2x RTX 4000 Ada cards, along with 8x modded RTX 4090s with 48 GB of VRAM. Running open-source models from Baidu and OpenAI, PewDiePie made a "council" of bots that voted on the best responses, and then built "The Swarm" for data collection that will become the foundation of his own model coming next month."
https://www.tomshardware.com/tech-industry/artificial-intell...
Yet another attention craving influencer who shilled crypto scams during the crypto bubble and is now marketing "AI councils" during the AI bubble.
He's not a serious or honest person, AI is just what he pivoted to after crypto. That's not innovation; it's attaching trendy branding to ideas that were already old when Marvin Minsky wrote The Society of Mind in 1986, three years before PewDiePie was zero years old in 1989.
The only thing PewDiePie's brought to the table is cleverly optimized YouTube thumbnails designed to attract clicks. The architecture is decades old; only his marketing and shilling is state of the art.
PewDiePie's Shady Promotion of Dlive (Cryptocurrency Scam)
https://www.youtube.com/watch?v=GYyNaQzZTo0
>If you want more concrete and in-depth explanations of what was touched upon in this video, here are the sources I used: Dlive ToS: https://community.dlive.tv/about/terms-of-service/ Dlive/Lino's Shady History: https://steempeak.com/@meno/an-objective-look-at-dlive-s-exi... Dlive overview: https://community.dlive.tv/about/welcome-letter/ 20 Million Chinese Investment: https://web.archive.org/web/20190716124653/https://www.coind... PayPal Suspension: • Is PewDiePie Supporting a Scam? Difference between Lino & Twitch Bits: • PewDiePie & Co-Founder Defending DLive? Is... Refusing to explain how they make profit: https://variety.com/2019/digital/news/pewdiepie-dlive-live-s...
Zonnetje in huis.
edit: also, 7 years on and after closure of the platform I cannot find evidence of rug pulls or exit scams associated with Dlive or any of their underlying crypto funding mechanisms, so before judging others on the content they consume maybe don’t take random YouTubers at their word and then don’t look into it 7 years after the fact.
I mean ... LLMs are sort of an extreme and living proof of linguistic determinism. Their behaviors are dictated almost entirely by disorganized language data, primarily English and Chinese. So you can't just add a language as native primary language in a quick post training, I think. There's no way that it would work.