The Washington Post burns its own archive
indignity.net
indignity.net
Paper books, vinyl audio records, old CDs, verified digital archives...
I was thinking it is worth creating a database of hashes for every human-generated piece of content, so authenticity can be checked. We already have the training data corpora, might as well ensure we know it is human-generated.
I have always been afraid that Google might be manipulating popular opinion, erasing, rewriting history by changing links it suggests. That is why I like archive.org; I can always go back to see state of web pages before. No rewriting history.
That is why I also keep my own repository of links https://github.com/rumca-js/RSS-Link-Database so that I always could return to links, that are important to me. I do not trust Google, or any big tech giant.
Back then there were microfiche scans. I assume they're not making new microfiche scans, but do they have scans of old newspapers available on a computer? If you want to do primary source research for a history term paper how do you do it?
Here's a list of newspaper archives containing such material:
https://en.wikipedia.org/wiki/Wikipedia:List_of_online_newsp...