Chemist Rafael Luque suspended without pay for thirteen years
english.elpais.com
english.elpais.com
I found this part quite impressive until I read the following part:
"Luque is constantly publishing papers. Last year he authored some 110 articles. So far this year he has published 58. The chemist admitted that since December, he has been using the artificial intelligence program ChatGPT to “polish” his texts. “These months have been quite productive, because there are articles that used to require two or three days and now I do them in one day,” he said."
Even if this is unrelated to the actual story, I find it a tiny bit disturbing. Is this the new standard practice? Let some robot spew out paragraphs of text to convey the scientist research?
The most prolific mathematician in history, Paul Erdős, published about 1500 papers over about 60 years. That comes out to a little more than 2 a month.
That's close to an upper bound on what should be possible without some form of cheating.
But he really did play a significant role in getting the results in every one of his papers.
> Magazinov mentioned that a non-existent “vegetative electron microscopy” appears in two studies by Luque published with Iranian colleagues.
How is that less concerning? Rewording conclusions or the abstract, even subtly, can change the meaning of those words and of those sections.
As long as the original researcher reads what's been written, and agrees with it, and the output gives an accurate description of the procedure and results, there's no harm done that I can see.
Edit: I mean in general, not necessarily in this specific case. This guy seems a little...questionable...for other reasons.
It's as different as ChatGPT is from an English-speaking scientist colleague.
For example, a colleague will probably ask if they're not sure which is the intended meaning of a phrase. ChatGPT will generate one of the possible meanings.
In all those cases, mistakes can be made and unintended meanings can creep in. The author has to check the output, definitely, but that doesn't mean those tools shouldn't be used at all.
...who were using even worse models to churn out papers about "vegetative electron microscopy."
While that was the issue that the university took with him, the extra logos on his lab coat, why does the university care about him? Why is he valuable?
Numbers. He boosts numbers with bullshit. Numbers that are misinterpreted to indicate value.
And he’s gotten good at bullshit. He’s even incorporated ChatGPT into his workflow.
So good at it that entities paid him to try and sneak some extra logos on his lab coat. He’s smart. He likely found ways to never accept money and yet fully utilize it for his benefit.
This all matters.
The system is fractally broken.
On the other hand this game has been played before. Everybody knows in Spanish speaking science that just adding one researcher from the Anglosphere with a nice English name to the list of authors is a seal of approval for many journals, and will open a lot of doors even if the researcher just agreed to sign on the paper.
It is entirely unsurprising that a non-native writer would do this. It would be more surprising in five years that anyone is not doing this.
From the article, the Spanish researcher gamed publication metrics by co-signing huge volumes of research papers, and proceeded to sell his affiliation to the highest bidder as it allowed low-tier institutions to game rankings and bump up their standing.
There's also another angle to his scheme, as his threat to the university of Cordoba consisted of "you har my interests and your university will drop in rankings".
The idea that someone can even read, edit and make co-author level contributions in papers aimed for a top-level journals, all in the span of 37 hours, is utterly ridiculous.
It's basically the equivalent of SEO spamming the Page rank algorithm with crummy pages with many back-links and repeated search terms. The search optimized term here is the author's name.
https://pubpeer.com/publications/7C9F0CCD493B1129135A3A918B0...
Search the page for the term and you'll see the relevant comment.
The term seems to stem from an OCR cluster-mishap, where a valid scientific document was once OCR'ed incorrectly. The page has 2 columns of text and the left column has a running sentence 'ending' with "vegetative" before continuing on the next line in the same column. But at the same height in the column on the right a sentence continues, visually starting with "electron microscopy".
The term is entirely meaningless in the field.
This isn't just "using a little ChatGPT to brush up my English" at all.
Other comments point out that the contents of abstract and article don't even match, the abstract being about bacteria but the article contents about gene delivery...
I don't need convincing that ML will have tremendous positive impact in math and science disciplines, it's obvious, but the way ML was used in this case is not a defensible usage.
It was almost as if many of the authors couldn't do the kind of math expected of high schoolers that apply to university to study science.
My takeaway was 1) to deeply distrust doctors as scientists (and as people who could think) and 2) mostly ignore the text surrounding the tables and graphs and just go straight to the data.
Please correct me if I'm wrong, but you're pointing out the absurdity of publishing every 2 to 3 days and implying that it was already essentiallly garbage he was producing, and chatgpt was just polishing the garbage (which IMHO, as you stated, the garbage production is clearly the bigger problem, not the polishing).
Yes, exactly.
If that first stat didn't pin your bullshit detectors into the red, you need new bullshit detectors.
It's really only when you let it introduce "facts" based on its body of training that it'll start going off the rails. If he's just running text through without checking the final results, that's really not wise, but the LLM is unlikely to start introducing novel facts or changing figures in response to a conservative request based entirely on the prompt text.
Using it as a writing aid the way he claims he is (not that I necessarily believe all his claims) is probably one of the more responsibles way to use GPT, honestly, because you have full opportunity and ability to vet the output. Hack together something that has all the right content, then make it readable with the AI engine the way you might with a human editor collaborating, while jointly making sure the editor doesn't change the actual facts of the content.
Whether or not he could be hacking anything of value together every 37 hours is another story.
I wouldn't be surprised if that statement wasn't a bold face lie, and Raphael Luque was just stapling his name to papers and proceeded with a clickfarm-like business model, where he sells the impact that his publication metrics has on institutions who request his services.
The defiant tone in his reply is evocative of other corruption cases where the criminal is so confident in his clever scheme that he even taunts defiantly everyone to challenge it. Even the old "I don't have a cent in my bank account" bullshit excuse is as old as time. Next he might just say that he only has good, generous friends who give him some presents from time to time.
>not just reading old things
ChatGPT is getting cited in comments, but this is the real lede.
This guy is extremely egotistical. And he's publishing papers with people he doesn't know. The article implies most of these papers are of low quality... I doubt he contributed in any way to many other than by adding his name. That's fairly common in academia even in normal times. But in these cases it's unlikely that he collaborated in any way with the authors and can't verify the content of the papers.
> Luque acknowledges that he skipped the established procedures to collaborate with other institutions, but attributes the sanction to envy and a lack of understanding.
Seems like they really lost a real gem with this one /s
He appears to have a very high opinion of himself.
Perfect application for this AI tool imo.
Yup - credential whoring and research fraud. It's the new frontier.
Using a machine to work more efficiently is not bad.
He wasn't sanctioned for using an LLM to speed up his work, so seeing this as the first reply could confuse some people.
I've been doing this for months now too. I can get docs done in a day that used to take a couple weeks because of constant interruptions.
Do you think Tesla or Einstein with a LLM model, would spit out 5 new discoveries per week instead of one every decade?
Who said the research was of flimsy quality?
Who said that the discoveries (as opposed to the BS padding journals require) was what LLM was used for?
They might or might not be. But not because he used an LLM to "polish them".
Something like physics, you can have a gap of months or years between a new theoretical breakthrough or the completion of a new machine giving you new raw data to churn through.
... but chemistry is something (no offense to chemists) you can do to the dirt outside your house, and every novel molecular arrangement can be paper-worthy.
If his was just (ab)using chatgpt to create the papers from thin air then it'd be pretty easy for somebody to show where the pre-chatgpt papers are and the post-chatgpt papers are by the drop in quality. However, nobody has been citing the content of his papers when they make the claims about quality so I doubt chatgpt has an impact on the quality of his papers.
Complaining about this is like complaining about the fact that people use IDEs to manage the imports in their Java class files. Ideally, you'd stop using a crap language that makes you wrap three lines of actual implementation in 50 lines of imports, declarations, and redundant-to-the-code comments to make doxygen (and a linter) happy, but if the institutional model prevents that... By all means use a good IDE to do half the work for you.
> The university has sanctioned Luque for working as a researcher at other centers, such as the King Saud University in Riyadh and the Peoples’ Friendship University of Russia in Moscow, despite holding a full-time publicly funded contract with the Spanish institution.
What's the problem here? That a non-native speaker now doesn't need as much time to write text in a foreign language?
Maybe if he would be a English, Oxford educated scholar, would not need ChatGPT. Could then maybe do a paper every 4 to 8 hours?
Is there a lot of overlap in the contribution that each of his individual papers make? Or are a lot of these replication studies, carried out by different teams in parallel, with him being a co-author on each of them? I'd be curious to read an academic chemist about how this is possible.
Generally possible everywhere, also done in the none-Google/Meta/... world of ML. Just look here: https://ieeexplore.ieee.org/xpl/conhome/9764877/proceeding - just a random collection of crap.
I'm reminded of a story doing the rounds about 25 years ago regarding an engineering document which had been translated by a machine. The team saw a reference to a 'water sheep' which they worked out meant 'hydraulic ram'.
I'm going to use GPT from now on, zero doubt.
It's just layman peeking at the Academia world and getting mad at a cloud.
Susana González, lost an EU grant of 1.8 million euro in 2017: https://www.nature.com/articles/ncomms14006
The wistleblower that disclosed that there were problems with the data and found images manipulated was fired and unable to find other work in research. Maybe is unemployed still, dunno. Somebody saved a lot of money also.
If it's the former, it really doesn't seem like a justified reaction.
[0]: https://en.wikipedia.org/wiki/Low-energy_electron_microscopy