TensorX
返回文献探索

Paper · arXiv 2310.20216

Does GPT-4 Pass the Turing Test?

Cameron Jones, Benjamin Bergen

17 upvotesOctober 31, 2023arXiv 预印本
AI 摘要

GPT-4 performed better than GPT-3.5 and ELIZA in a public online Turing Test but failed to reach human-level performance, highlighting the importance of linguistic style and socio-emotional traits in human detection.

GPT-4Turing TestGPT-3.5ELIZAlinguistic stylesocio-emotional traitsnaturalistic communicationdeception

Abstract

We evaluated GPT-4 in a public online Turing Test. The best-performing GPT-4 prompt passed in 41% of games, outperforming baselines set by ELIZA (27%) and GPT-3.5 (14%), but falling short of chance and the baseline set by human participants (63%). Participants' decisions were based mainly on linguistic style (35%) and socio-emotional traits (27%), supporting the idea that intelligence is not sufficient to pass the Turing Test. Participants' demographics, including education and familiarity with LLMs, did not predict detection rate, suggesting that even those who understand systems deeply and interact with them frequently may be susceptible to deception. Despite known limitations as a test of intelligence, we argue that the Turing Test continues to be relevant as an assessment of naturalistic communication and deception. AI models with the ability to masquerade as humans could have widespread societal consequences, and we analyse the effectiveness of different strategies and criteria for judging humanlikeness.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Does GPT-4 Pass the Turing Test? | TensorX