PaSh: Light-Touch Data-Parallel Shell Processing
arxiv.org
arxiv.org
The joys of peer review.
It's the academic contribution to the wider body of knowledge reviewers will focus on as the paper, not the code, will be the thing presented at conference.
Code will likely be spaghetti. Academic research is often more about "finding new stuff" than "building robust stuff".
Doesn't always make sense IMO, but that's the way of the world.
Simple things like grep can be split and parallelized on a single system with a ssd. The pain comes back to doing something like combining results.
Additionally, commands have so many options they can move between simple to parallelize, to more complicated.
I also wonder if the approach should be to build a query planning layer al la pash, or addressing the parallelism in the command itself. I.e their sort example.