Is this why they bought Mendeley, a document-management start-up who did a lot of data mining on academic papers, including terabytes of Elsevier's?
Hint: probably. I used to work there.
Hint: probably. I used to work there.
EDIT: found it! Great!
Zotero's use of Google Scholar to extract metadata from pdfs made it easy when starting with several thousand pdfs.
But the NS editor Was/Is a bit of a Luddite so nothing came of it.