“Close to 10% of the papers we receive show some sign of academic misconduct”
retractionwatch.wordpress.com
retractionwatch.wordpress.com
To make matters worse, the same experimental setup is frequently used to run experiments that result in several publications. When you have several papers authored by different overlapping groups of researchers in various stages of review simultaneously things can get very confusing. The first paper you submit is frequently the last one published. The last paper you submit will then wind up referencing arxiv pre-prints instead of published articles, and if you did plagiarize yourself by mistake, the publisher of the paper submitted last might have absolutely no chance of detecting it. On top of it all, there is frequently interference from editors and reviewers. A well-written and coherent paper on a theory an experiment testing that theory might be considered "too long" or "too confusing" by a referee, and the editor will jump on that comment and demand the paper be split into two, sometimes only one of which they feel like publishing in their journal. If the researcher had done this of his/her own initiative many would consider it CV-padding!
Yes, self-plagiarism can be a means to pump up your publication list, but it also happens innocently because publishing is such a confusing counter-intuitive mess! It's probably an idiotic idea, but some kind of pan-journal version control system, even just for figures, would be a tremendous headache reducer if done right!
For fun, trying googling "It increases the amount of crap", and you will find that your very own post meets the academic definition of plagiarism. Both for containing a single plagiarised phrase and percentage wise (46% plagiarised!).
I've seen post grads spend weeks using software to look for identical phrases of 4 words in their final work and previous work and then slightly rephrasing a sentence to avoid duplication of a common phrase and the dreaded self-plagiarism. Then running the thing again and inevitably finding the new wording was a duplication as well. I couldn't see it as anything other than a colossal waste of time. If you are writing about the same topic it's inevitable that you will phrase things in the same way at times.
All of this does nothing to combat the CV padding you worry about because it only concerns reusing the same exact language, not talking about the same ideas over and over in different ways.
As an outsider it seems that academia has descended into a kind of madness about plagiarism and self plagiarism. I think this is largely thanks to software companies trying to sell their plagiarism detection systems.
Plus, if the competition wasn't so gruesome and that evaluation criteria weren't shallow, the pressure to pad one's CV wouldn't be so great.
"Holy ethical violation batman, the slaves are slipping out of their manacles and they certified on their honor they would not do that"
I know it's an oversimplification, but the essence of it is true: curing cancer is already hard enough. It shouldn't be made any harder by ambitiously exceeding your dead tree quota.
Sauce: abandoned an academic career :-)
I wouldn't minimize the problematic quality of having a bunch of crap in research or what-all else. But the thing is that, as you know, the academic world puts extreme pressure on people to both follow the rules and to makes themselves look good. When any violation of a rule-set can be called an "ethical violation" when it's really not that, it cheapens the whole concept of ethics. Which isn't surprising given how monumentally unethical it is to set up system.
Essentially, I can't help seeing people with a lot of power engaging in seriously unethetical things and happy to blur the lines between that and people cuts corners in response to pressure (and that stuff winds up being more pressure on people which doesn't actually stop the corner-cutting).
I don't feel uneasy because I'm doing anything remotely unethical---I'm not. I feel uneasy because it sounded a bit like the description in this article was over-broadly defining self-plagarism.
I will say that academic (i.e., scientific) conferences should never include "recycled" or "old but interesting" material. There has to be some new research results. Otherwise, it's not science. The point of these conferences is not to entertain, it's to push forward the scope of knowledge.
Either way, I agree with everyone else in this thread as well. The system of publish or perish is 1. dumb 2. old 3. not moving science forward. It's become more of a "I need xyz papers published this year" instead of "I'm curious about xyz and how it solves abc". That isn't science, it's a career/ego metric.
Good news is that there are groups (like myself, science exchange, figshare etc) who are working to fix that. It's a matter of time, but a social change to science is starting.
There is so much pressure on researchers to publish. I really do think the way to solve this sort of problem is to find ways to give researchers credit for other forms of contributions.
This is what figshare is doing with datasets, and what we're doing with peer-review: http://blog.publons.com/post/61380784056/announcing-doi-supp...
Either production-quality open source code, or pedagogical code. I'm looking at the Stanford Pintos kernel and MIT xv6 kernel. While there were minor papers from those projects, I think they were more like a labor of love. When you consider the coding effort, those projects probably took 10x the effort than a typical paper.
But yeah it would be better if a little more time was spent on code vs. papers.
I actually attended a talk from an Adobe researcher talking about software abstractions some years ago. He advocated that you should be able to get a Ph.D. for finding a good abstraction, e.g. for say modeling a paint brush or something. There are lots of bad ways to write code but only a few good ones. Even better would be to write it in a way so that other people can actually learn from it.
As a PhD student implementor, yes, my theory colleagues get way more publications than I hypothetically could even if I were a better student than I am. That makes it hard to get an academic job, but otherwise doesn't matter too much.
A PhD in CS does not mean "extremely good software engineer," it means "scientist."
Your second argument is circular -- I'm saying a valid research goal should be to write solid and useful code, and then explain it.
For example, you could write an OS kernel or kernel subsystem to fill some particular part of the design space. A great but rare example is what the authors of Lua have done.
Academia wants new ideas. Production (industry) generally revolves around making a new idea useful to a broad range of people, by which point it's an old idea. You can take something written in academia and put it into production, but if you do that in your Ph.D. it almost certainly isn't helping you to finish.
I'm saying it would be better for society if the academic culture emphasized the craft of coding, rather than solely "new ideas". The whole point of this article is that the emphasis on "new ideas" incentivizes fraud.
Academic culture changes faster than you think. I expect that the structural changes caused by online courses will have a big effect in the near future.
On the other hand, maybe they really are trying to improve themselves and I am being too judgemental.