Yeah I dont think a single current LLM would fool me in a turing test - I would obiously use all kinds of prompt injection techniques, ask about 'dangerous' or controversial topics, ask about random niche facts in varied fields, etc.
Thats a good point actually, I hadn't thought about deploying hacks and jailbreaks in a turing test but thats exactly what should be done, if its being done adversarially