For example, this is why the archive obeys robots.txt policies retrospectively. I can't access my personal web pages from 20+ years because at some point the university changed it's robots.txt policy. (The domain no longer exists, freezing that policy quasi-permanently.) But at least they're archived. Maybe at some point before I die I'll get to see them again, but at least it's a possibility. More importantly, my "sacrifice" (such as it is) is a small price to pay to help ensure that archive.org preserved access to material that would have been cut off if websites felt that they didn't have control of availability once archived.
We can't expect archive.org to fight our battles for us, whether regarding censorship, copyright overreach, or other issues. It's great that they do to the extent that they do, but their success is partly a consequence of them picking their fights and emphasizing their character of being a quiet, harmless archiver. It would be counter-productive to ask them to risk that credibility and power. And it's entirely unnecessary.
What surprised me was that you say they respect the robots.txt, but the data is archived anyway, so they are crawling anyway, but only publish according the rules defined in the robots.txt?
If so that is news to me.