This reminds me of books. I’m sure the majority of books from over a hundred years ago are lost because they weren’t popular. We haven’t really noticed their absence…
Especially if you include independently published books that weren't widely circulated. I wonder what percentage of total books this is.
My grandfather published a book before he passed away. It was never sold online or in any big retail stores. Once the last hard copy is lost, it's gone forever.
i believe the Library of Congress will archive that book for you if you mail them a hard copy. assuming it has a ISBN, you’re in the US, etc.
Not all books in the state libraries are equal. Historical copies and popular authors (popular among researchers, a much bigger set already) are exhibited and get attention, John Doe's book of family recipes gets sent to some giant dark warehouse people rarely visit.
It is easy to forget that it is an 18th century solution born from 18th century approach to knowledge. Back then, bibliographies of everything printed in certain year in certain country could be compiled, and they were supposed to be more than just lists, to help other men of books keep up with Progress.
So, 10 hours of white noise yes, some person's personal blog where they poured their heart out, no.
Beautiful.
Kierkegaard, Thoreau, Dickenson and Melville, for instance.
If their works had been lost, "we" probably wouldn't have noticed any of those absences either.
If they were in one of the university libraries that Google scanned, they're not "lost." But you're right; you can't read them. Congress should mandate that the Library of Congress, at least, get a copy to preserve them for the ages.
Read the Atlantic article
https://www.theatlantic.com/technology/archive/2017/04/the-t...
for the sad story.
For example, if LLM NN model weights are distributed with IPFS instead of corporate infrastructure (basically zero redundancy) the popular models would be very available, and have essentially near zero chance of being lost.
To state that again, the llama models likely have tens of thousands of downloads, which would mean tens of thousands of servers and backups of the data, versus what we have now, which is essentially just one.
We need IPFS for data distribution. Tightly knit integration with git repos is an obvious match as well.
Slack’s limited search history is a feature too - forces you to document using appropriate tools and not endless email threads…
I'm sure that I don't know what you mean. My employer is on a paid plan for Slack, and searches cover everything, as far back as I wish to go. Are you thinking of the limitations on the free license?
The other great thing about slack is that anyone from company can start an account without IT’s approval…