What is your justification for AI training not being theft?
It does not assume that.
Otherwise there's obviously a legally relevant distinction between a human mind which is ascribed agency to decide if and how to use its memories of copyrighted material, and importing into an information retrieval system which can't help but spit out transformations of parts of its inputs on demand, (including lossy representations of the Getty watermark if it's fed enough Getty material, or an exact facsimile of an image if that's all it's trained on...)
That's a perfect oxymoron.
Copyright uses infringement, which is not theft: it's non-rivalrous, and it contains a number of exceptions.