Bot Submissions to Comment Website Can’t Be Distinguished from Human Submissions
techscience.org
techscience.org
> To quickly weed out inappropriate comments, I handpick from generated comments those that ensure a high coherence and high relevance sample for submission.
So basically it's a validation of GPT-2 making sense with small amounts of text. Judging from the demo test page, they are pretty good texts, but he said it himself that larger texts betray the bot. So, i m not sure what he's trying to prove by using MTurkers, since this does not attack the problem mentioned in his introduction: the fake FCC comments were weeded out through text analysis, not via human work.
In all, i'm not sure if this is something that people didn't know about gpt-2. The title is certainly not justified, perhaps "Curated bot comments can't be distinguished by humans to be obviously fake" would be better, but also more banal.
Think https://www.reddit.com/r/SubSimulatorGPT2/ is more impressive than a study where half of GPT-2 comments handpicked for being human-like by one human were accepted by another human. Particularly given that some of the comments in question were three or four words long...
It's a very good idea to make sure submissions come from humans, but it's also slightly overstated and alarmist to state that model-generated text is reliably passing the Turing test.
For my part, I got only 60% (12 out of 20). This may be because English is not my first language.
Compared to a wiki-style website, where all angles of the argument can be collected into one place to make a cohesive comparative overview; as forum-users, we are left stranded in noisy content, and we rely on making heuristic judgements based on popularity of certain opinions and stubbornness of certain commenters. Bots make easy work of exploiting these flawed heuristics.
All media has its flaws, and I still prefer to check forums for the greatest diversity of opinions. Strangely, I have noticed an unintuitive aspect of forums: smaller forums appear to have a greater diversity in opinion than larger ones.
A lot of news is now generated by bots, Bloomberg itself has 30% of its content almost entirely generated [0], so does that render said news "Deepfake news"?
Or is it only when we're attempting to be alarmist?
[0]https://www.nytimes.com/2019/02/05/business/media/artificial...
What would make you believe that? If the content generated is of decent quality, timely and accurate (as of the time), why would people skip it?
It doesn’t have to be a human or bot but a human and bot together :)