I don't see why a code review couldn't be the same.
I don't see why a code review couldn't be the same.
From your description, it sounds like you presume the software is accurate.
I’d expect there to be peer reviewed studies about its error rate, a description of how it works, limitations, etc. For instance, if you ask for a one sentence factual response to a question, there’s no way it will be able to distinguish LLM from human. I’m sure it has other limitations.
I met someone that built one of the better packages in the “don’t copy your answers from google” space a while back.
He routinely gets pleas for help from people his software has falsely accused of plagiarism (the tool is open source and popular and maintained by others).
He generally sends strongly worded letters to educators explaining the limitations of the tool, why it’s probably wrong in this case, and that says he will happily testify against them in court as an expert witness.
I imagine the commercial vendors are less transparent, but not more accurate.
> I'm sorry, Dave. I'm afraid I can't do that.
Subjectively, if a student can knowledge-guide an LLM to generate response from a specific perspective, I think it is very debatable whether such usecase should be outright punished.
For example, giving it a small corpus of my previous handwritten work and asking it to respond in the same style is exceedingly easy and very, very hard to detect. Especially if you run a small algo on it afterwards to introduce your category of typos etc.
It is only a matter of time before services will pop up for students that automate/wrap this.