And I'm sure students will use tools to have every other paragraph written in the style of a different AI, in an attempt to defeat this fingerprinting.
I saw that someone already created a Github repo for a Python script that strips the watermark out of Claude generated text. It was released, I think, within 24 hours of the announcement. I cannot attest to how well it works, but I found it humorous nevertheless.
Yeah, wait for LLM "scrambles" that put every paragraph and then the whole text through multiple re-write/edit style cycles.
How would that change anything? The proposed watermark is applied while the output tokens are being chosen, taking that text and running it through an LLM again would just repeat the process.
It says it only applies to passages over 200 words, so you could use a different LLM to rephrase every other paragraph and undermine the watermarking.
You take the output and run it through another LLM with "please re-write this in xxx style". Then you repeat that a few times on different part of text and glue it all together at the end.