Any way to shift the text processing to the client? I'd like to use the bookmarklet on some academic papers (many behind a paywall) and the few I've tried only seem to parse the abstract...I assume this is because the text processing is happening server-side, but I could be wrong.
Alternatively, could you release your backend code as well? I'd like to run this on larger corpora.
Very elegant and useful project!