Presumably “scraped” isnt the right term here. They already have the raw data, they
Won’t be “scraping “ it from the website they’ll just be investing it from where they store it
I’d probably say Meta trained their models using all self-hosted, public AU citizen’s data.
But it doesn’t really sound as scary as “scraped” to non-technical users.
Perhaps "consumed" would work just as well and be more accurate.
In some legislations there are rules about scraping. And for many less technical people it sounds scary.