Caltech Announces Open Access Policy
caltech.edu
caltech.edu
The decision of our faculty to make their papers freely
accessible online will ensure that the global community of
researchers, students, and casual followers of science and
engineering will learn about our work at earlier stages, enabling
them to put it to use for the benefit of society.
I love how they mention "casual followers of science". To me this is a huge deal. I'm not a member of academia, but I really enjoy reading papers from a wide variety of fields (in the spirit of "learn everything about something and something about everything"). I never intend to stop, and I know lots of "laymen" who do the same.The unavailability of papers, lack of centralized tools, and terrible search interfaces have been incredibly frustrating. I can't wait for a day when we get a centralized, non-profit, publish/subscribe consumer service for all the papers that ever get published by major research universities (with a good search tool). The value of such a service to the public and society would be enormous. As more and more people get educated and used to dealing with science, this could be as big a deal as wikipedia.
We're not quite there yet, but this is a huge step by CalTech in the right direction.
Me too. This is quite beautiful. After I left academia, subscribing to all the journals I used to have access to through my university became economically unfeasible, for obvious reasons.
I work at a growing OA Publisher (www.ubiquitypress.com ;) and can completely emphasize - it's one of a bag of problems we're looking at. The search interfaces on individual journals is the standard mixed affair however there are many indexes out there that aggregate articles or just their metadata with decent navigation. Some even specialize in OA: https://www.google.com/search?q=OA+index
I'm not too unhappy with Google Scholar lately. It seems to have nearly complete coverage of anything that can be found in an online paper archive (both institutional archives and journals' own archives), and a better search interface than those archives do. Via extracting references it even indexes a large amount of older material that's not yet digitized; of course it can't link you PDFs in that case, but it still provides the citation and can return offline papers by title/author/date in search results.
It's a great tool, and I use it all the time, but it's certainly no replacement to a properly-managed bibliographic database.
I still applaud what they're doing. If the research is even remotely funded by taxpayers, then everyone should have access to it. (I consider it partially funded if the school gets NSF grants, or is a non-profit)
Two examples come to mind: first developers of distributed databases might want to follow the progress on distributed data-structures such as CRDTs [1]. Apparently Riak developers at Basho are fast at implementing such new theoretical results into their product [2].
[1] http://pagesperso-systeme.lip6.fr/Marc.Shapiro/papers/RR-695... [2] http://vimeo.com/43903960
I am pretty sure that Basho as an organization is not a subscriber to any pay-walled CS journals. So having such articles as Open Access helps make it discoverable by the engineering community, just by googling, sharing links over social networks, mailing lists or news sites such as HN and reddit.
Similarly developers of (big or small) data analytics tools might want to follow the research on machine learning algorithms so as to implement the state of the art and empirically evaluate it against the previous baselines. As an engineer contributing to the scikit-learn project this is basically my job. I am no researcher myself but I (and the many other scikit-learn developers) try to help transfer the results from academic research to make it available to our users community that reaches out of the traditional academic circles.
Having researchers publish their results in Open Access venues reduces the friction to transfer new results to practical, productionalized implementations (e.g. as open source projects maintained over the years). Open Access is one of the tools to help break the traditional Researcher / Engineering boundary as agile development tools and the devops movement helped break the Developer / Sysadmin boundary. To me Open Access is a fundamental building block of "Agile Science".
You probably know about this already, but if you're interested in anything biomedical in nature, the US National Library of Medicine's PubMed interface to the MEDLINE database (also managed by the NLM) is exactly what you describe: http://www.ncbi.nlm.nih.gov/pubmed
MEDLINE is a bibliographic database that indexes all of the abstracts going back to the 1960s from literally every major (and almost all minor) journals that do anything even remotely related to biomedicine, and also include a surprising number of physics and CS journals. It features a robust and easy-to-use API, and has since the 1990s. All MEDLINE entries include obvious bibliographic metadata (titles, authors, abstracts, dates, etc.) but also include human-curated index headings, which makes searching way easier. PubMed's interface makes it easy to construct complicated and accurate queries, and the NLM offers online access to their reference librarians, as well.
Furthermore, any published research that is funded by the US National Institutes of Health has to have its full text (including figures, references, etc.) posted to PubMed Central (PMC), which is another online database that the NLM maintains. This is an actual law- if you get money from the NIH, your publications have to end up in PMC following a short embargo period- I think it's about six months post-publication?
All of the PMC content is available free of charge, and it also has an API. It is not a complete full-text mirror of MEDLINE, because not everything in MEDLINE was NIH-funded, but for research going back over the last 7-10 years or so it is often surprisingly complete.
This is one area where biomedicine is way ahead of computer science. As somebody who's worked a lot in both fields, I definitely find myself missing PubMed whenever I need to look for CS literature...
As an aside, the National Library of Medicine is truly one of the hidden jewels of the US government. Its budget is tiny, but it somehow manages to produce fabulous work and performs a vital service for the research community. And almost nobody has heard of it!
If not, most paper will probably not be deposited...
Best part:
>Faculty may still grant exclusive rights to their publishers, either permanently or for an embargoed period, but to do so, they must request a waiver from the open-access policy. At other institutions with open-access policies, such as MIT and Harvard, faculty have requested waivers for about 5 percent of the total number of papers produced, usually to comply with the requirement of a few publishers that want a formal waiver in order to even consider manuscripts for publication.
5%, and in those cases, because of dinosaur publishers. This is a very conclusive sign for the shift in mentality towards open access.
On the flip side, it is a point of friction, if small. It did give me a very specific moment for my complaint to be registered (even if overridden).
Wow! Really? I am a math professor, and in mathematics I have never heard of this happening. It would strike me as professional suicide on the publishers' behalf: holding a knife to the neck of their golden goose.
It seems that I was not alone in this perception:
http://academia.stackexchange.com/questions/9958/why-do-univ...
Although I am in 100% in support of Caltech's policy, I would have guessed that the issue was mostly theoretical and that they were only taking a stand on principle. Either Caltech is bluffing, or I stand corrected. I'm not sure which. But in either case, kudos to them.
I do wonder, how common is such a policy at other institutions? I assume this must be uncommon, since it's apparently newsworthy?
UCL are exceptional in this regard and we work closely with the institution and many of it's students and alumni publishing their (OA) books and journal articles: http://www.ubiquitypress.com/ebooks and to download these books: http://www.doabooks.org/doab?func=publisher&pId=1194
OA shares many parallels with the rise of Free/Open Source software its uncanny.
I sure wish those lectures had been recorded and I could refresh my memory on the more interesting ones.
I've also seen some of the MIT online videos, and my recollection is the Caltech lectures I attended were of comparable quality.
This is the one caveat. It would be great to out these journals. Seems likely is public/501c money going into the research on these, too.
At this point I'm left wondering:
1. Is the typeset version that much of a value add? Do people not realize how relatively easy it is to do these things on apart from publishing houses?
2. Is reading a single column version all that bad? We do it on the internet every day.
I only wish Aaron Swartz had lived to see it.