I submitted a meme from November and asked it to explain it and it seems to be able to explain it.
Unfortunately chat links with images aren't supported yet, so the image:
the response:
The humor in the image arises from the exaggerated number of minutes (1,300,000) spent listening to “that one blonde lady,” which is an indirect and humorous way of referring to a specific artist without naming them. It plays on the annual Spotify Wrapped feature, which tells users their most-listened-to artists and songs. The exaggeration and the vague description add to the comedic effect.
and I grabbed the meme from:
https://later.com/blog/trending-memes/
Using the human word "understanding" is liable to set some people off, so I won't claim that ChatGPT-4 understands humor, but it does seem possible that it will be able to explain what the next meme is, though I'd want some human review before it pulls a Tay on us.
and I'm in a bad mood now seeing how unfunny most of those are
here's the next one from that list:
the response:
The humor stems from the contrast between the caption and the person’s expression. The caption “Me after being asked to ‘throw together’ more content” is juxtaposed with the person’s tired and somewhat defeated look, suggesting reluctance or exhaustion with the task, which many can relate to. It’s funny because it captures a common feeling of frustration or resignation in a relatable way.
Interestingly, when asked who that was, it couldn't tell me.
“human word” as opposed to what other kind of word?
I don't think you need to improve that much current LLM so they can detect actual harm threats or hate speech from any other type of communication. And I think those should be the only sort of banned speech.
And if facebook wants to impose additional censorship rules, then it should at least clearly list them, and make the moderator AI explain what are the violated rules, and give the possibility to appeal in case it is doing wrong.
Any other type of bot moderation should be unacceptable.
Example: Picture of a plate of cookies. Obese person: “I would kill for that right now”.
Comment flagged. Obviously the person was being sarcastic but if you just took the words at face value, it’s the most negative sentiment score you could probably have. To kill something. Moderation bots do a good job of detecting the comment but a pretty poor job of detecting its meaning. At least current moderation models. Only Meta knows what’s cooking in the oven to tackle it. I’m sure they are working on it with their models.
I would like a more robust appeal process. Like bot flags, you appeal, appeal bot runs it through a more thorough model, upholds the flag, you appeal, a human or “more advanced AI” would then really detect whether it’s a joke sentiment, sarcasm, or you have a history of violent posts and it was justified.