For example if there was an ad hominem attack "Fuck you Ellen Pao." The auto moderator would allow the post but would highlight the text red with a small bubble identifying it as ad hominem, users could react appropriately.
Or, maybe you have a long well thought out argument that has a straw man in it. The auto-mod could highlight the straw man to point out that it exists and maybe even format the post to show the point from which the rest of the position is based on that.
Maybe this wouldn't have to be actively presented but just a tool that comments were run through so users could be warned prior to their submissions. Obviously the complications with this would be ridiculous and you could never stop training the algorithm that was processing your text. But the implications of having such a thing seem pretty incredible.