Peepeth.com is experimenting with features like that. Currently, posts get a toxicity score from a Google API before they're even posted, and users are asked to confirm if it's above a certain threshold. Lots of false positives, but it seems to be reducing profanity.
Disclaimer: I'm the founder.