Internet Archive generally allow websites to control if they get indexed/mirrored or not, via the robots.txt, so websites can decide for themselves.
Luckily, we have other grassroots movements like ArchiveTeam that doesn't care and archives anything deemed valuable to be archived, website owners be damned.