1. Legal loopholes given OpenAI's advertising aspirations and model training needs
2. Data retention and rising threats of fascism that historically have not served the persecuted very well when fascist regimes get access to said data
3. Risk from centralized collection of that data with a company whose software I do not control in a world where enshitification and lock-in is the norm.
I really wish OpenAI did more to espouse exactly this: "When I say no training, I mean no training. No gimmicks around data vs derived data, synthetic data, preference data, etc." and ideally provide technical reasurrances that this is impossible (eg: certain technical ZDR approaches, etc.).
Do you happen to have a favorite reference to point me at that would document some of those official assurances to the nuanced detail we've discussed here?