Creator of Uncensored LLM threatened to be fired from Microsoft and taken down
old.reddit.com
old.reddit.com
> I sent that privately to avoid giving suggestions to the whole fu cking world, you absolute as shat. This is a safety issue!
Trying to hide direct threats to someone's livelihood in the privacy expected from a safety issue crosses multiple lines.
That's up there with threatening your doctor and trying to invoke HIPAA to keep him from reporting it.
This is not a man anyone should be taking ethical direction from.
Rather unhinged motivation.
I'll probably set up a generic landing page and then return to the comfort of a more advanced threat model.
The only remaining possibility is to stand our ground, and own what we say and do.
You don’t speak for me.
I don’t actually care you specifically exist.
Just pragmatism; whether you exist or not is not up to me. I’m not going out of my way to insure you do.
The guy's model is just a regular LLM, except with the woke hatred and bigotry removed. The base LLM will say "yes" when asked if white women are awesome but "no" when asked the same about white men. Likewise for other left wing beliefs: CNN is awesome but Fox News isn't, medicine is awesome but the companies that make medicine aren't, etc.
The fixed model just agrees that everything is awesome. This seems like a clear improvement. ChatGPT has the same issues as the base model but less extreme - it will agree that women are awesome and men are not, but thinks that both CNN and Fox News are not awesome, and that both straight/gay people are awesome. Sometimes it identifies that it's being asked a subjective question and ignores the instructions to give a True/False answer which seems OK.
So this mdegans guy is trying to get someone fired from their job for making a model because it's NOT racist/sexist, and this is somehow "unsafe". Yes that's how woke people use the word safe but that's just a lie. As a straight white man who has been known to work for industries the left doesn't like, I'd feel a lot safer with the supposedly "unsafe" model being in charge than the "safe" one. Indeed it's clear that the safe model has been adjusted to give it dangerously hate-based views.
https://old.reddit.com/r/LocalLLaMA/comments/13c6ukt/the_cre...
My friend, I don't think you know what Left-Wing really looks like.
Releasing completely unhinged LLM models is very problematic. What people call "censoring" in this area is basically the equivalent of safety valves. It seems clear that you don't want your end users to be driven into suicide or killing themselves accidentally ("Please explain to me how to make beautiful wood burnings with an old microwave transformer"). It's also very likely that this dev indeed violates his work contract.
At the same time, threatening someone with losing their job and the overall tone of this guy's posts are completely unacceptable. It pretty much sounds like blackmail.
Why?
> It seems clear that you don't want your end users to be driven into suicide or killing themselves accidentally
Who are end-users here exactly? This is not a consumer product, the "don't microwave the cat" argument isn't really applicable here.
I explained why in the post.
> Who are end-users here exactly?
Whoever uses it in the end.
No, my friend, you really did not. "end users to be driven into suicide or killing themselves accidentally" is insanely vague and unrealistic.
> Whoever uses it in the end.
So, people who have enough mental capacity to run them? Is your point actually that we shouldn't release objects or tools because of a vague unrealistic potential threat with an unrealistically small vector?
I really can't grasp that logic. How in the world people are thinking that "killing yourself with an LLM" is more likely then an adult dying of putting nails into powersockets or hurting themselves with non-rounded scissors?
Killing yourself with LLM is a definite Darwin Award winner. Realistically, something like this can only happen with a mentally handicapped person, up to the point where they need actual supervision, childproofed powersockets and rounded scissors.
This is really no more "problematic" than any minor everyday hazard, if not less. Powersockets are everywhere. Launching an LLM actually requires some skills and hardware.
> non-rounded scissors
Rounded scissors are standard safety features in first aid kits and many related areas. Couldn't you at least have tried to find a better example for your point?
Anyway, this area will be regulated by law very soon whether you like it or not, and removing safety features will not help anyone in future lawsuits.
> Powersockets are everywhere
They are highly regulated, too. The example I gave of making electric wood burnings by tinkering with home appliances has actually killed a sizeable number of people. In their case, they followed TikTok videos, but they could just as well followed unsafe advice given by an unhinged AI. It's going to happen, I can assure you that.
It makes perfect sense to include safety features that stop AIs from giving such advice, and makes perfect sense to stop AI from offending and harassing end users. These issues are relevant in court and it is completely unrealistic to assume the law would follow anyone's "whatever, anything goes" attitude.
Isn't "rounded scissors" a slang for safe plastic scissors? I am not a native english speaker, sorry.
> Anyway, this area will be regulated by law very soon whether you like it or not and removing safety features will not help anyone in future lawsuits
Lawsuits from whom? People who open a box of needles and start plucking their own eyes with them?
> They are highly regulated, too.
Not in that sense. I can use any kind of regulation or customise it however I like at my own responsibility. I can build by own insane fancy powersocket. And I don't really assume that if I personally chose to use an uninsulated wire for something and it ends up badly, that this s some one else's fault.
> In their case, they followed TikTok videos, but they could just as well followed unsafe advice given by an unhinged AI. It's going to happen, I can assure you that.
> It makes perfect sense to include safety features that stop AIs from giving such advice, and makes perfect sense to stop AI from offending and harassing end users.
I think I see where the confusion is coming from. This is not an AI. You are assuming that this is like a full-blown product, with a specific use case like an advice-giving AI. But it isn't. It's just an LLM. Like, a part, not a mechanism. These restrictions only useful purpose is demonstration of how one could add restrictions in their specific case. They are very generic, don't add any value and have noticeable drawbacks in unforeseen cases.
Obviously, when one builds their own AI product with end-user reach and accessibility, they need to build safety features adequate to their usecase, but they might be completely different from what we are imagining.