She Was Falsely Accused of Cheating with AI – and She Won’t Be the Last
rollingstone.com
rollingstone.com
> Stivers received an email from her professor indicating that a portion of it had been flagged by Turnitin as AI-written
> Last month, a Texas professor incorrectly used ChatGPT in trying to assess whether students had completed an assignment using that software. It claimed to have written every essay he fed in — so he temporarily withheld a whole class’ final grades.
> William Quarterman, a senior and history major at the college. A professor had run his exam answers through an AI detection called GPTZero, which returned a positive result. The professor gave Quarterman a failing grade and referred him to the same student affairs office
Let's let professors indict students based on untested (and amateur) AI algorithms that not even the vendors completely comprehend and subject the accused to a process in which they are presumed guilty until they prove their innocence-- while they mortgage their future paying 5 to 6 figures for the privilege.
These are the real AI "safety" issues that need to be addressed. The only "danger" posed by AI is in its total lack of accountability. This is outright abusive behavior on the part of colleges, who wasted no time in weaponizing it to conduct virtual arbitration. The irony of an AI system flagging students for copying what someone else wrote is apparently lost on everyone.
Lol, happens everywhere, power of memes and all that. Only difference is you know more about ChatGPT than them.
Turnitin advertises the feature as "98% accurate". That's a pretty terrible accuracy rate for potentially destroying someone's life.
The utter hypocrisy of having to write "ethics / impact" blurbs for class projects, when the school can seemingly implement this sort of life-changing technology without a long trial period and carefully designed guard rails around how to interpret the results is galling.
The best case scenario is that there are so many false positives that they get recognized for the garbage they are, but I think in reality there may be a small population of students who are unlucky enough to write in a way that gets flagged significantly more often.
How do you defend yourself against that sort of situation? "Well, your classmates didn't get flagged. Why did you?"
The problem with this technology is that there is no concrete evidence that you can use to prove / disprove the accusation. There is no single piece of plagiarized content that you can dig up and point to as the smoking gun.
The ethics of letting this technology go live is breathtakingly stupid. I'm concerned this sort of laziness is a harbinger of things to come, and it fills me with dread.
Assuming the 98% applies to both false negatives and positives:
For a given 20k person school, assuming they each write only one paper a year that's run through this, an administrator gets to expel 400 students a year!
What fun!
The more insidious issues is you have a sub-population of students that write in the "wrong" way. That means the average student might never get a strike, but this unlucky sub-population might get multiple strikes. That would seem to point to the results being "correct" -- after all, you'd expect a cheater to cheat multiple times.
If this population skew is significant enough -- say, 1% of the population writes in the "wrong" way, and the multiplier effect is 10x as strong for them -- you might never get a flag for a normal student. But what you're measuring is not cheating, but deviation from the norm in writing style.
Memory that pops up is Barton Fink being told to write a script for a wrestling film.
Also sort of topical in high school had a quiz on a short story that I had not read. But I had read some of the authors other short stories. So I started out with 'While I failed to read 'story title'. Based on the author and the title I think the story is about bla bla. And the teacher took off 2 pts for not reading the story first.
Haven't these people all taken statistics course? Or at least enough of them. Just to read that one number and do some rough estimations of what it means in any meaning...
very low chance, unless you are one of the two people this week.
Even if the number was correct, which it ain’t, this is insanely bad
I used to be a teacher and even when I strongly suspected a student of cheating, if I couldn't prove it, I wouldn't even mention it to them or anyone else. Weak evidence doesn't mean partial guilt.
In general I found those tools mostly useless and very easy to fool by making minor changes. I had better luck putting the original writing prompts into chatGPT, saving them all to a text file, and cross referencing when dealing with student work that didn't pass a sniff test.
All of this is the kind of detective work that most graduate students are not qualified to perform easily.
A better check is that there’ll be a mismatch between content and citations, or the citations will be entirely hallucinated. As far as I can tell ChatGPT is incapable of generating actual citations.
As for checking hallucinations I definitely do not have the time to go that deep on 75 submissions a week. If that could be automated it'd be helpful. The students are correct that I can afford to spend only about 5 minutes per essay on grading (they're short, 200-250 words).
plug in your finished work, making up whatever stats and arguments and assertions suits your taste or agenda, and CiteGPT will back-solve matching snippets that support your writing, inserting all in line citations and building a reference list automatically. Maybe even suggest minor tweaks to make your arguments match the available supporting material.
It does mean that good writers may be able to get away with cheating, but then they really didn’t need the practice to begin with.
I think the more reasonable solution is to just wait for the university to take active measures against it, and accept that this is just a new development in how students cheat, which they have always done. The fact that the administrators are dumping this all on teachers and TAs is totally unfair to the actual workers.
At a certain point, telling if someone "cheated" on a research paper by evaluating only the text of the paper because ineffective because a student can always pay someone else to do the assignment for them. Instead of devising better assessment techniques, they do the same old things.
I've always learned a thing or two after completing a research paper, but ultimately they always felt like busy work. The goal was always the paper, not an actual objective that requires research.
I wonder if a tiny upside could be that we stop teaching students to write like that. The formulaic, vapid mechanisms we teach them aren't merely unpleasant to read. They're opposite to the goal of writing, which is to get a point across. When we teach them to write like Mad Libs, they're not really learning anything about communication.
There’s no reliable way to identify LLM generated content. Even if OpenAI could watermark someone could download Alpaca and run it locally.
That kind of decentralized legal action would get costly for TurnItIn very quickly.