Nice. I found you can beat this by picking the word least likely to be selected by a language model, because it seems like the alternative choices are generated by an LLM. “Pick the outlier” is the best strategy.
This is presumably also a simply strategy for detecting AI content in general - see how many “high temperature” choices it makes.