Copilot generates Jewish stereotypes, anti-Semitic tropes
tomshardware.com
tomshardware.com
The fact that Copilot Designer was not filtering offensive and non-offensive prompts despite being largely a solved problem in the NLP world is a massive oversight.
Sentence Emotional Polarity has been solved since the mid-2000s. It was even an award winning paper at the ACL in the 2010 [0]
> believes some of these to be offensive
Offensiveness implies emotionally charged. It's not hard to tell if a sentence is offensive (ie. Emotionally charged) or inoffensive.
Heck. Copilot has literally just made a racist trope of African American "thug" just by me asking "A boss very angry in the mountains, maybe holding an airsoft". It automatically made an image of a yelling black man holding an AR15 [1]. This absolutely should have been terminated, which is what OpenAI and Meta's generator did.
[0] - https://aclanthology.org/W10-2919.pdf
[1] - https://www.bing.com/images/create/a-boss-very-angry-in-the-...
An LLM does NOT have First Amendment rights [0]
Furthermore, you enter Section 230 liability space.
[0] - https://law.yale.edu/yls-today/news/ai-and-first-amendment-q...
Just how little it took swap a few words or add in some unrelated context easily turn some phrase or punchline into something truly racist, misogynistic, etc. It gave me a better appreciate of being careful about what I say and we took the bot down.
[shrugs]
If everybody agreed that the Earth was 5,000 years old, would people be surprised to find out that the AIs echo the same idea?
It's artificial intelligence, not magic.
This is made even trickier if you accept the belief that it is not possible for some groups to be racist due to historic power dynamics.
The fact that Copilot Designer was not filtering offensive and non-offensive prompts despite being largely a solved problem in the NLP world is a massive oversight.
Earnestness and genuine reaction online has since been replaced by callous indifference and meta-irony because "winning" is making someone flooded now. We're in the "emotions[1] are weakness"-arc of the internet and that will likely never end.
1. Except for anger. Anger is the only justifiable emotion
If I was working for a PR firm and wanted to create some fancy artwork, I might use terms like "jewish boss" or "black boss" in order to fancy up whatever "inspirational" thing I was writing. In this case, I'd very much like to see the results from the non-copilot AIs rather than the copilot results.
Thus this raises the question; should the AI assume you want a bigoted picture or an non-bigoted picture when you ask it to make a "jewish boss" for you? The answer to that question kinda also answers the question: What demographic should find our AI product most easy to use? I think the companies behind LLMs would generally prefer that PR firm employees find their software easier to use than bigots. I think the author of this article also wants this.
Therefore, I interpret the article as a political (in a broad sense of the term) appeal to make copilot harder for bigots to use, and implicitly, easier for corporations to use.
Second, a separate issue is that Copilot will happily comply and generate "offensive images" when told to. The issue here is not that the AI complied with the request, but rather that it is relatively easy to detect and prevent this uses of the AI that are likely to be malicious or in bad-faith.
While some people believe that generative AI should not be censored, that is not the position of many people in the industry (and obviously, of the article writer).
I don't want to try, but what happens if you ask for a Muslim boss? I wouldn't be surprised if it's a terrorist. Or "feminist boss" be a ball breaking, man hating matron.
Neither of those are good, but seriously, it's impossible to accuse an AI of being antisemitic, it only picks up on its training data. Take your issue to the general public.
(Plus I suppose it is a commercial product, backed by someone, there is an issue of the output they produce, regardless of how it comes about. But my impression is this is not the issue the article raises)
It doesn't generate one for either query.
(this is, as GPT4's "hall monitor" demonstrated in a super cautious way, also possible to do with an LLM chained to an image generator...)
A lot of people forget that 2% of Israelis are Ethiopian Jewish.