This is no stretch of the imagination due to existence of fraudulent copyright takedown requests being submitted daily to Youtube against legitimate channels (with Youtube giving said legitimate channels very real copyright strikes against their accounts regularly, with only very visible cases having a chance of reversal)...
https://arxiv.org/abs/2202.05262 https://arxiv.org/abs/2210.07229
9:55 LoRA
14:45 Hypernetworks
user input ---> | prompt filter | GPT | output filter | --> sanitized output
1) Block Personally Identifiable info to get into training data : Maybe have a redact such info before passing it to model . Just having a layer between I/O and the processing .
2) Have a sort of VCS for models and rewind and retrain if such is needed to be done . ( Seems quite unfeasible )
3) If it is not an info but rather an opinion , train the model to counter biases until a netural or agreed stance is reached
OpenAI doesn't remove data from the trained models, they filter it at the output level. They also remove it from training data, but of course the model lives on.
I'm a strong supporter of a "do not encode" header / metadata on content that allows individuals, content creators and providers to tell AIs not to encode specific text / images / documents in the first place.
Sounds like there is a decent story that can be written about the right-to-forget in GPT/LLM era ... similar to "The Eternal Sunshine of the Spotless Mind" for human memory.
"Hallucinations" will be a main story-telling device in this one as well.
Can you craft a specific prompt to get it to hallucinate info about someone non-famous without mentioning them in the prompt?