English
Related papers

Related papers: People cannot distinguish GPT-4 from a human in a …

200 papers

We evaluated GPT-4 in a public online Turing test. The best-performing GPT-4 prompt passed in 49.7% of games, outperforming ELIZA (22%) and GPT-3.5 (20%), but falling short of the baseline set by human participants (66%). Participants'…

Artificial Intelligence · Computer Science 2024-04-23 Cameron R. Jones , Benjamin K. Bergen

We evaluated 4 systems (ELIZA, GPT-4o, LLaMa-3.1-405B, and GPT-4.5) in two randomised, controlled, and pre-registered Turing tests on independent populations. Participants had 5 minute conversations simultaneously with another human…

Computation and Language · Computer Science 2025-04-01 Cameron R. Jones , Benjamin K. Bergen

Everyday AI detection requires differentiating between people and AI in informal, online conversations. In many cases, people will not interact directly with AI systems but instead read conversations between AI systems and other people. We…

Human-Computer Interaction · Computer Science 2024-07-15 Ishika Rathi , Sydney Taylor , Benjamin K. Bergen , Cameron R. Jones

The Turing Test, first proposed by Alan Turing in 1950, has historically served as a benchmark for evaluating artificial intelligence (AI). However, since the release of ELIZA in 1966, and particularly with recent advancements in large…

Human-Computer Interaction · Computer Science 2025-05-06 Avraham Rahimov , Orel Zamler , Amos Azaria

We present "Human or Not?", an online game inspired by the Turing test, that measures the capability of AI chatbots to mimic humans in dialog, and of humans to tell bots from other humans. Over the course of a month, the game was played by…

Artificial Intelligence · Computer Science 2023-06-01 Daniel Jannai , Amos Meron , Barak Lenz , Yoav Levine , Yoav Shoham

This paper critically examines the recent publication "ChatGPT-4 in the Turing Test" by Restrepo Echavarr\'ia (2025), challenging its central claims regarding the absence of minimally serious test implementations and the conclusion that…

Artificial Intelligence · Computer Science 2025-12-30 Marco Giunti

Large Language Models based on transformer algorithms have revolutionized Artificial Intelligence by enabling verbal interaction with machines akin to human conversation. These AI agents have surpassed the Turing Test, achieving confusion…

The current cycle of hype and anxiety concerning the benefits and risks to human society of Artificial Intelligence is fuelled, not only by the increasing use of generative AI and other AI tools by the general public, but also by claims…

Human-Computer Interaction · Computer Science 2025-01-30 Sharon Temtsin , Diane Proudfoot , David Kaber , Christoph Bartneck

We examine whether a leading AI system GPT4 understands text as well as humans do, first using a well-established standardized test of discourse comprehension. On this test, GPT4 performs slightly, but not statistically significantly,…

Computation and Language · Computer Science 2025-01-22 Thomas R. Shultz , Jamie M. Wise , Ardavan Salehi Nobandegani

Advances in artificial intelligence (AI) raise important questions about whether people view moral evaluations by AI systems similarly to human-generated moral evaluations. We conducted a modified Moral Turing Test (m-MTT), inspired by…

The pursuit of artificial intelligence has long been associated to the the challenge of effectively measuring intelligence. Even if the Turing Test was introduced as a means of assessing a system intelligence, its relevance and application…

Robotics · Computer Science 2025-07-23 Lavinia Hriscu , Alberto Sanfeliu , Anais Garrell

The Turing Test is no longer adequate for distinguishing human and machine intelligence. With advanced artificial intelligence systems already passing the original Turing Test and contributing to serious ethical and environmental concerns,…

Machine Learning · Computer Science 2026-01-01 Adam Winchell

As AI becomes increasingly embedded in daily life, ascertaining whether an agent is human is critical. We systematically benchmark AI's ability to imitate humans in three language tasks (image captioning, word association, conversation) and…

In his seminal paper ``Computing Machinery and Intelligence'', Alan Turing introduced the ``imitation game'' as part of exploring the concept of machine intelligence. The Turing Test has since been the subject of much analysis, debate,…

Artificial Intelligence · Computer Science 2023-08-03 David Harel , Assaf Marron

The world has seen the emergence of machines based on pretrained models, transformers, also known as generative artificial intelligences for their ability to produce various types of content, including text, images, audio, and synthetic…

Artificial Intelligence · Computer Science 2024-10-10 Bernardo Gonçalves

We administer a Turing Test to AI Chatbots. We examine how Chatbots behave in a suite of classic behavioral games that are designed to elicit characteristics such as trust, fairness, risk-aversion, cooperation, \textit{etc.}, as well as how…

Artificial Intelligence · Computer Science 2024-01-02 Qiaozhu Mei , Yutong Xie , Walter Yuan , Matthew O. Jackson

The role-play ability of Large Language Models (LLMs) has emerged as a popular research direction. However, existing studies focus on imitating well-known public figures or fictional characters, overlooking the potential for simulating…

Computation and Language · Computer Science 2024-04-23 Man Tik Ng , Hui Tung Tse , Jen-tse Huang , Jingjing Li , Wenxuan Wang , Michael R. Lyu

The Turing test examines whether AIs exhibit human-like behaviour in natural language conversations. The traditional setting limits each participant to one message at a time and requires constant human participation. This fails to reflect a…

Computation and Language · Computer Science 2025-05-30 Weiqi Wu , Hongqiu Wu , Hai Zhao

The development of advanced generative chat models, such as ChatGPT, has raised questions about the potential consciousness of these tools and the extent of their general artificial intelligence. ChatGPT consistent avoidance of passing the…

General Literature · Computer Science 2023-04-26 Arend Hintze

The pursuit of human-like conversational agents has long been guided by the Turing test. For modern speech-to-speech (S2S) systems, a critical yet unanswered question is whether they can converse like humans. To tackle this, we conduct the…

Artificial Intelligence · Computer Science 2026-03-03 Xiang Li , Jiabao Gao , Sipei Lin , Xuan Zhou , Chi Zhang , Bo Cheng , Jiale Han , Benyou Wang
‹ Prev 1 2 3 10 Next ›