I have no idea how you came to that conclusion. Unless your training pipeline involves actively querying one of Anthropic models, no they can't. And if it does you're distilling their model.
I don't even think they can believe it themselves, it's in reality they are just trying to throw fear, uncertainty and doubt about potentially cheaper offerings.
Not what that means.
Crocodile tears "is a colloquial term used to describe a false, insincere display of emotion" [1]. Defending yourself against an attack vector you just exploited is between savvy and hypocritical.
The fun part is that you will never know if your neural net classification project is getting silently sabotaged because their classifier doesn't work!
With this in mind, I don't want model to be proactively instructed and encouraged to sabotage without telling me.