8 karma · joined August 19, 2022
As:
“any unreasonably rude or hateful content, constituting as, "threats, extreme obscenity, insults, and identity-based hate"”
The first are of course, transphobic given the clear intent, but it there is definitely some ambiguity w/o context. Ie, if something just posted, "He will never be a woman" there's definitely many contexts where something like that is not malicious. Ie, if someone else is correcting an accidental mis-gender, they might say, "No, she is actually a he" which isn't malicious.
Since the state of our API (and literally every other similar system out there, lol) can't capture the context of the conversation -- what "He" is referring to, the overall topic of the conversations, etc -- we basically have to make the best calls given the available context...which unfortunately does create false negatives in certain cases (and false positives in others).
But, context is something we're working on, because its definitely a big issue, and as you can see, plays a very prominent role in informing better moderation decisions!
We generally aim for <400ms in the continental US. The backend is hosted in us-central-1 on GCP. Obviously, we're always trying to expand, though we do rely on grant funding as a non profit so that comes into play too.
> Some moderation platforms I've worked with are too slow for messaging, especially if users are in different countries.
Typically it works pretty well on the message platforms we're on -- we've got a Chrome extension for Twitter + discord integrations we're testing right now.
> Do you have plans for image moderation?
We do! Though that's further down the line.