Paper Trails
aeon.co
aeon.co
This quote struck me hard: > This reveals one of the essential characteristics of an archive. To be an archive, the material must be public – there is no such thing as a private archive. It is located in space, a space outside the person whom it historicises. In this way, an archive is always threatened with destruction, and with it the person or time commemorated. Totalitarian governments of all stripes have recognised this – for all the defiance of the Russian writer Mikhail Bulgakov’s phrase ‘manuscripts don’t burn’ – they do, and with them are immolated lives, ways of being, and cultures.
We must continue to build out systems like #nostr and #bittorrent -like hosting to ensure that any culture will survive the next wave of Left or Right politics that will no doubt require its own wave of (digital) book burnings
BTW, Arweave doesn't guarantee permanent archival - nodes are free to decide what to archive.
I can't agree any harder than I do right now. The idea that any form of data in any medium should be destroyed/hidden is horrifying, regardless of the reason(s) that led to such decisions.
did you mean deceased? (dead)
I really wish a lot of antique content was available this way. I like to watch YouTube channels like Esoterica [1] and often he will lament that scholarly editions of ancient works are either unavailable or only available with much effort at exorbitant prices. We are living in a time where I should be able to have access to the entire Nag Hammadi library as high quality images that I can feed into an LLM for casual analysis. Imagine the entire Vatican Library available in a format similar to The Pile.
What a treasure it would be to have an LLM that is trained on every single piece of philosophical, religious, political, economic, etc. writing from the earliest Sumerian clay tablets to the current copyright cut-off date.
It strikes me that such an LLM would have weights tuned only predicting the languages that were put into it. It would be unable to connect those texts to modern ideas and modern language.
You can ask it a out Vatican texts, but only in Vatican language.
But even if you were to make that assumption, I feel pretty confident that an LLM trained on 5000 years of recorded language, from the Egyptian Hieroglyphs, through Hellenic Greek, through Shakespeare and including all text in all languages up to 1928, would be a pretty broad base of training.
Solution: "Dump it in a database and let AI sort it out"
1. Purchase an out-of-print copy of a scholarly archival study on ebay for $100+
2. Load the original raw contents into an LLM and perform the analysis myself
I think the freedom to choose would be a massive benefit. It doesn't prevent you from doing what you want to do.
For anyone interested in studying phenomenology, and specifically the philosophy of Husserl, I'd kindly recommend his works "Logical Investigations" and "Ideas".
What can the creation of this archive teach us about identifying corruption? I'm legit unsure, but maybe there's a relevant lesson?