In a dual-chat, five-minute Turing test variant, 71% of judges correctly identified a Llama 3.2 1B chatbot as AI, versus only 44% in a simple single-chat version with prompt engineering.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.HC 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
The Turing Test Is More Relevant Than Ever
In a dual-chat, five-minute Turing test variant, 71% of judges correctly identified a Llama 3.2 1B chatbot as AI, versus only 44% in a simple single-chat version with prompt engineering.