I think filtering on upvotes/downvotes, comments/views, sub, user and whatever metrics they have on the content can help AI companies train on somewhat reasonable things. Blend it with Wikipedia, scientific papers, reliable newspapers and you're golden?
Metadata is what makes gold out of poo, I assume model developers can "train negatively" too if metadata suggests they should.