I'm not sure what they'd get from training on that
I had a very basic React question about useState while porting some vanilla code last week which all models of all stripes I've tried it on have been confidently and completely incorrect about, up to stating the code absolutely will not work, even when I take a turn to assert that I ran it and it does, so there's plenty of shit in there already.
If human response is "That's BS", "fuck off", or something similar, mark as bad assistant message.
If human response is "huh" or "cool", mark as good assistant message.
If on ChatGPT, watch how much scrolling user does. If there's a lot, its somewhat likely that the LLM outputted something useful.
That strategy would have holes of course but as long as its better than guessing something like that would be a useful heuristic.
Even very weak human signals can be immensely valuable over large enough datasets.
Marking is not a trivial task though. Use some AI system to mark it and you get a 99.something% filter maybe but whatever that remainder is leaks through. Over time your filter may get worse as a result.
Grok is the only one that swore back at me. I kinda liked that. The others are way too polite, "Artificial Intelligence? Artificial Canadians, more like", my uni-going kid joked.
In Gemini you can turn off Gemini Apps Activity (warning: deletes your chat log, you need to copy paste everything into notes)
Highly recommended.
The real process involves submitting a request on another one of OpenAI's sites and awaiting a confirmation email (either their privacy or platform site).
Feel deceived and violated? Yeah, you, me and millions of other people, welcome to the club.
"I previously opted out of model training by writing to the support team. Will you continue to honor my opt-out?
Yes. If you opted out by contacting support or using our privacy form, your account will represent that request."
https://help.openai.com/en/articles/7730893-data-controls-fa...
I thought it boiled down to credibility.
Apple - alleged Siri eavesdropping: $95M [0]
LinkedIn - alleged unauthorized ai training on private messages: ?? [1]
Google - alleged unlawful data collection in Texas: $1.4B [2]
[0] https://www.usatoday.com/story/tech/2025/05/11/apple-siri-95...[1] https://www.itpro.com/security/privacy/linkedin-faces-lawsui...
[2] https://www.businessinsider.com/google-alphabet-settlement-t...