* Stack Exchange, but only one of their sites: superuser.com
* Wikimedia: wikimedia.org, wikidata.org, and wiktionary.org (but not wikipedia.org)
* Image hosting: imgur.com, giphy.com, tenor.com, deviantart.com, artstation.com
* Audio and video hosting: soundcloud.com, vimeo.com, dailymotion.com
* Blogging and web hosting: livejournal.com, squarespace.com, substack.com, typepad.com
* Fan fiction: tvtropes.org, archiveofourown.org, fanfiction.net
* Miscellaneous indie web properties: bogleheads.org, craigslist.org, ifixit.com, instructables.com, knowyourmeme.com, mozilla.org
Overall this is a very strange list.
I think it's a proof of concept on how we can theoretically make searches better. Anyways, I'm pushing an update removing all the sites you mentioned. Do you have any other suggestions?
The implementation details take time to get right. And getting this far in an afternoon as a non-developer is pretty respectable.
If you want to have the most reach / impact with something like this though, it will take consistent, sustained effort over a long period of time, and it will be a thankless job, at least for awhile.
Assuming the dump is accurate, this is at best an impractical and at worst a misguided blacklist.