Saving 25,000 Manuals
ascii.textfiles.com
ascii.textfiles.com
Some of these old books have some gems of information in them that you just won't find in modern books. Things like hardware schematics are commonplace in old programming manuals. In the 8-bit era it was usual to cover everything from assembling the computer to programming it in the same book. There weren’t dozens of blogs posts and “for dummies” books either: the manual was it. It had to cover everything.
Then there are the useless factoids, my favourite of which is that Tron is actually a command from old versions of Basic: TR(ace) ON. There is a corresponding but less cool-sounding TROFF command. I was rather disappointed that the names in Tron Legacy were all random nonsense, presumably because nobody involved realised that that the names in the original were all real computing terms from that era.
I've never really used Google Groups, but then this might be a good thing in this case. Just trying it quickly, there doesn't seem to be any way to subscribe to a group without getting a Google account, I guess after doing so you might be able to just forward it all to whatever email you want. But to an outsider it certainly doesn't look like "an email listserv", and it's not entirely clear how to use it as such.
Whether this makes Google Groups "shitty", I'm not sure. I guess a large part of it is about expectations. Does Google still claim to "do no evil"? I'm sure if it was Microsoft that was running the service then no one would be surprised.
(Regarding the actual interface, atleast I find for example gmane to have a much better interface, Google Groups seems very bloated with lots of dead space. One additional thing, as a none native English speaker, Google Groups shows for each message, a rather large box with a question asking if I'd like a translation of that post. I have a hard time believing this feature is used enough to warrant the amount of space that's spent on it.)
Microsoft had a bunch of newsgroups running onicrosoft servers. Some of those newsgroups propagated outside MS.
It was a nice experience. There was the Microsoft MVP (most valued professional) programme where people with in-depth knowledge of a product and reasonable people-skills would provide peer support. Including the term MVP in a web search would return the websites for those people. For years (and probably still) those pages were sources of xcellent information about MS products.
Here's one example. Search terms [excel fill handle MVP]
http://dmcritchie.mvps.org/excel/fillhand.htm
This page loads instantly. It tells me what the fill handle is, what it does, how to use it, how to trouble shoot it, how to get more information about it, how to use the existing MS Excel documentation to get this information. It covers some advanced fill handle use, and some gotchas. It alsotells you whih parts of this information are not found in the help files.
In this case: MS did good[1] things for Usenet. Google fucked it.
[1] ignoring the massive amounts of sub-optimality caused by OE having some Usenet functionality but being buggy and insecure.
Unfortunately, "Do No Evil" does not include, "Don't Be A Dick."
Google has killed 15-20 years worth of information that's not only of interest for it's history, but for the still useful conversations on damned near any topic you can think of.
Ok, so not everything is still useful. I can't argue that since there were plenty of sports and entertainment groups, but there were also many groups whose information doesn't really become stale. In hiding this data, Google has fucked over anyone who might have benefited from finding it.
Hell, the very damned first time I ever heard of Google was an announcement on USENET. I thought they were pretty slick then and I started using Google instead of DogPile(? I think it was dogpile. That was a long time ago.).
Some history: Google acquired DejaNews and the contents of other usenet archives, and has largely let all of that data languish, with what can sometimes be very weak search abilities of the archives via Google Groups (no hits for XYZ in an active group for XYZ, for instance), and where Google doesn't make the Usenet archives available and visible via the main search Google engine, and has generally become somewhat of problem.
Then there are the folks that dredge up a decades-old usenet thread — possibly having no idea what Usenet is — and post to it, and with the usual hilarity that ensues. "Hey, is PDQ still available?" to a post offering PDQ that originally posted in 1997, etc.
By some appearances, Google Groups is headed in the same direction as Google Reader.
One great thing is that USENET posts were dated, so you could look for the first occurrence of a word, phrase, rumor, etc. That's hard to do on the web.
Then, a few years ago, the search function started to get worse with every release, with reduced functionality and more and more unreliable results. You might search for a term and see a result from 2005. Then, when you limited results to before 2006, you would get nothing! Around the time they were trying to force Google-plus integration, it became nearly unusable.
There used to be an advanced search page[1], but that's gone. When you search within a group, there's a pull-down for advanced search options, but that's missing from the main cross-group search. It seems that search operators[2] may still work.
The same kind of degradation happened with Google News Archive and, to a lesser extent, Google Books.
This is because it fails to follow its own specification for "escaped_fragment" URLs.
http://developers.google.com/webmasters/ajax-crawling/docs/s...
http://groups.google.com/robots.txt
Google wants webmasters to "opt in" and let Google access #! URLs without using Javascript (they want to be able to use a bot). This is done by providing an alternative "escaped_fragment" URL.
But Google themselves will not allow users to access Google Groups #! URLs without using Javascript. They will not provide alternative "escaped_fragment" URLs. Why not?
Search for the newsgroup comp.unix.soures and you will see an error message that Googlebot was unable to access Google Groups because of its robots.txt
Of course, you will also see comp.unix.sources has (fortunately) been mirrored on numerous servers elsewhere, probably well before the Google acquisition of Deja.
I think Google should give these old newsgroups back to the community in their original format.
Edit: Nevermind, searchable database for which manual you want to buy a hardcopy of, not the actual manual...
Dear Googlers (er, Alphabeters?) and others, please consider making a tax-deductible donation to the archive:
Mel (from The Story of Mel) was identified thanks to an old LGP-30 manual. I'd hate to miss out on treasures like that in the future.
After poring over the photographs for a few minutes it appears that quite a few of these are electronics manuals. But really, I have no idea what is special about this collection or why it is suddenly so important to save it at the last moment - presumably I'm just supposed to just go along with Jason Scott's intuition that it's important.
Am I missing something?
I will make a new entry mentioning you especially.
https://en.wikipedia.org/wiki/Jason_Scott
"In January 2009, he formed "Archive Team",[12] a group dedicated to preserving the historical record of websites that close down. "
"Jason Scott was hired by the Internet Archive in 2011"
Or his about page:
http://ascii.textfiles.com/about
"In 1998 he started a website called TEXTFILES.COM whose original mission was to make available the thousands of BBS textfiles he’d collected in his youth, but which has now expanded greatly in all directions of computer history."
Recommended video:
https://www.youtube.com/watch?v=Gq70QKa7588
"That Awesome Time I Was Sued for Two Billion Dollars"
Any plans to scan them and put them on line?
It doesn't give the price and I imagine it's quite expensive though.
There's also https://youtu.be/dByUFS-YJDo?t=1m34s
and https://www.youtube.com/watch?v=uX0g4aNynro&feature=youtu.be...
None look terrible cheap and they all seem to need a human to load up the book to begin with though they can turn pages automatically.
The Usenet etext / ebook groups used to have some useful faqs about best scanning / OCR techniques. I don't know the right words to make Google return anything useful.
The only thing not automated is turning the pages. Books come in a wide range of sizes and materials, they may be damaged or fragile, etc. Human hands are dexterous and self-repairing. I believe they did experiment with automated page turning at one point, though I don't know the details.
In some ways it's getting worse, really; at least you don't need proprietary hardware and software to read a 50 year-old book.
I think General Radio got their start in the 'teens of the 20th century.
I also have old tools and instruments from 50+ years ago that I have no information on. Would not be surprised if there were old Sheffield Instruments manuals somewhere in there.
Of course we also had vacuum cleaners, steam engines, automobiles, combines, and all sorts of other things that needed manuals. At https://youtu.be/rUg_ukBwsyo?t=310 you can hear mention of the manual that came with a 1925 steam car. ("In the handbook it says 'Things for your man to do every day.'")
To be concrete, here's a copy of the instruction manual for a 51-P Crosley radio receiver from 1924: http://www.crosleyradios.com/pdf/51P_Operating_Instructions.... . A "Manual of wireless telegraphy (radio) for the use of naval electricians" from 1915: http://babel.hathitrust.org/cgi/pt?id=nyp.33433023245677;vie... . And "A practical guide for the use of the Edison phonograph" from 1892: http://babel.hathitrust.org/cgi/pt?id=nyp.33433023245677;vie...
(actually a curious question. Why is this interesting to you?)
People looking for manuals on old computers and devices combing for that wisdom (or just curiosity) should check out this site too:
Someone should look into http://store.diybookscanner.org and just scanning them all, then processing them with a pdf->text software. and just have them searchable so that they can be used in the future if they need to be...
Whats the time limit for archiving these types of documents in terms of copyrights?
I think manuals like this should all end up on bitsavers.org