I mostly agree, however, there are some issues I'm foreseeing. I've done some NLP work myself and have had access to very private information because of it. The team we were working with were all highly educated and well paid. Because of this, we were aware of the implications of the data we were working with, so at one point, in one of our randomly selected data sets, we found data that was from a public figure and removed it to avoid any possibility of a conflict of interest.
Where my concern is, is that a lot of annotating work is now being done by low wage, low education workers and because of the increasing demand of annotated data sets this group is increasing. Also, because this work is increasingly getting outsourced, there is less direct control of who is doing the work and therefore the chance that a bad apple slips through is bigger. That's what's scaring me going forward.