Mozilla, GitHub and Figshare team up to fix the citation of code in academia
thenextweb.com
thenextweb.com
Yeah... well there is the achilles heel of this whole thing. "Always" in internet time means "2 or 3 years, or until we fail to close another round of funding" in real-world time. I'm probably being a cynical old fogey again, but let's see in 5 years time if this whole thing still exists (mozilla, for example, doesn't have a very solid track record of keeping projects up, to put it mildly) before I start putting URLs to this in papers that people will probably still reference every once in a while 5 years from now.
Then again, if everybody thinks like I do, it'll never get off the ground - classic catch-22.
Figshare on the other hand, not so much.
http://figshare.com/blog/Ensuring%20persistence%20on%20figsh...
Twelve research universities keep copies of all the data and will make it publicly available should Figshare implode.
Other data repositories, like Dryad, do the same thing, and grant agencies with data deposition requirements usually require you to deposit data somewhere that has a long-term plan to ensure access.
So as long as the twelve universities manage to update all the DOIs to point at the new locations, Figshare DOIs should be reliable.
That's why some scientists are going to use Google's Compute Engine. They just need the machines. The researchers have their own C++ and Python scripts. They can live with some complexity. They are happy with HTCondor which is awesome for running big computational jobs.
Sharing data results with the world is awesome. But transiting to a new platform is again a big problem.
Let's say you release your project to GitHub & figshare and now have a DOI in hand. What are you supposed to do with it? Do you ask your users to cite this DOI if they use your software? If so, what text should accompany the citation? How do you track citations to your code? Will they show up in Google Scholar, Scopus, Web of Science?
And what if the journal one of your users is submitting to doesn't accept figshare / github citations? It's unfortunate but true that many publishers disallow citations to unpublished / non-academic works. This is why many scientific software projects have resorted to publishing papers on their software — it's a hack to make a software project fit into the traditional social system of scientific credit.
DOIs are a technical glue that binds together the thousands of academic publishing outlets, but they do not solve the scientific or cultural issue of what is the minimum viable citable scientific product, and how those citations are generated, propagated, or valued.
Securing a DOI only solves a small slice of the problem of scientific credit — a point most colorfully expressed by this blog post from CrossRef, the largest DOI registrar for academic work: http://crosstech.crossref.org/2013/09/dois-unambiguously-and...
Worst case, GitHub and Figshare both go under and we're back to where we started. The one hesitance I have is about the Figshare/DOI'd repo being frozen in time - I keep making arguments to myself about how this is a good idea or a bad idea.
Is that what they're doing? That actually seems like a pretty good idea, if it is, I hope they are! And if they're not, it would be trivial to do.
Github's UI still makes it easy to see what happened after (or before) that point (including the 'latest' version), but if you're citing software used as a tool for research results, it makes sense to be able to cite the actual software that really was used, not it's hypothetical future evolution.
cat .git/refs/heads/master
Why are three large-ish organizations feel it necessary to combine their powers for something this trivial?