Did people find it to be as challenging when you showed it to them as some of us are here? Did you expect that level of complexity?
Did people find it to be as challenging when you showed it to them as some of us are here? Did you expect that level of complexity?
Yes, I think people definitely find it challenging. I'm keeping track of the correct and total guesses for each snippet, right now people are at almost exactly 50% accuracy:
correct | total | pct
---------+-------+-----
6529 | 12963 | 50EDIT another poster guessed GPT2 each time and found the frequency was 80 percent
SELECT id, code, real FROM code ORDER BY random() LIMIT 1
Unusual to be this lopsided (1-in-7), but not crazy.
[0] https://en.wikipedia.org/wiki/Misuse_of_p-values#Clarificati...
It's rare that a p-value is what you want, but for answering "how unusual is this case", it's the exact right tool for the job.
What I found more worrying is how terrible some of the real code was :D