English
Related papers

Related papers: People cannot distinguish GPT-4 from a human in a …

200 papers

Human intuition has been simulated by several research projects using artificial intelligence techniques. Most of these algorithms or models lack the ability to handle complications or diversions. Moreover, they also do not explain the…

Artificial Intelligence · Computer Science 2011-06-30 Jitesh Dundas , David Chik

As a result of continuing advances in computer capabilities, it is becoming increasingly difficult to distinguish between humans and computers in the digital world. We propose using the fundamental human ability to distinguish between…

Cryptography and Security · Computer Science 2019-07-09 Nasser Mohammed Al-Fannah

Many psychophysical studies are dedicated to the evaluation of the human gestalt detection on dot or Gabor patterns, and to model its dependence on the pattern and background parameters. Nevertheless, even for these constrained percepts,…

Computer Vision and Pattern Recognition · Computer Science 2018-05-28 José Lezama , Samy Blusseau , Jean-Michel Morel , Gregory Randall , Rafael Grompone von Gioi

This paper examines how individuals perceive the credibility of content originating from human authors versus content generated by large language models, like the GPT language model family that powers ChatGPT, in different user interface…

Human-Computer Interaction · Computer Science 2023-09-07 Martin Huschens , Martin Briesch , Dominik Sobania , Franz Rothlauf

Recent claims suggest that large language models (LMs) underperform humans in comprehending minimally complex English statements (Dentella et al., 2024). Here, we revisit those findings and argue that human performance was overestimated,…

Computation and Language · Computer Science 2025-05-15 Adele E Goldberg , Supantho Rakshit , Jennifer Hu , Kyle Mahowald

In this paper we apply our understanding of the radical enactivist agenda to the classic AI-hard problem of Natural Language Understanding. When Turing devised his famous test the assumption was that a computer could use language and the…

Computation and Language · Computer Science 2022-04-26 Peter Wallis

People are known to judge artificial intelligence using a utilitarian moral philosophy and humans using a moral philosophy emphasizing perceived intentions. But why do people judge humans and machines differently? Psychology suggests that…

Computers and Society · Computer Science 2023-09-20 Jingling Zhang , Jane Conway , César A. Hidalgo

Human-like agents are an increasingly important topic in games and beyond. Believable non-player characters enhance the gaming experience by improving immersion and providing entertainment. They also offer players the opportunity to engage…

Artificial Intelligence · Computer Science 2025-06-11 Maciej Swiechowski , Dominik Slezak

Increase in computational scale and fine-tuning has seen a dramatic improvement in the quality of outputs of large language models (LLMs) like GPT. Given that both GPT-3 and GPT-4 were trained on large quantities of human-generated text, we…

Artificial Intelligence · Computer Science 2023-03-31 Philipp Koralus , Vincent Wang-Maścianica

As dialogue systems and chatbots increasingly integrate into everyday interactions, the need for efficient and accurate evaluation methods becomes paramount. This study explores the comparative performance of human and AI assessments across…

Computation and Language · Computer Science 2024-09-11 Ike Ebubechukwu , Johane Takeuchi , Antonello Ceravola , Frank Joublin

Background: The increasing deployment of Conversational Artificial Intelligence (CAI) in mental health interventions necessitates an evaluation of their efficacy in rectifying cognitive biases and recognizing affect in human-AI…

Computers and Society · Computer Science 2025-02-11 Marcin Rządeczka , Anna Sterna , Julia Stolińska , Paulina Kaczyńska , Marcin Moskalewicz

In this pilot study, we investigate the use of GPT4 to assist in the peer-review process. Our key hypothesis was that GPT-generated reviews could achieve comparable helpfulness to human reviewers. By comparing reviews generated by both…

Human-Computer Interaction · Computer Science 2023-07-13 Zachary Robertson

In the post-Turing era, evaluating large language models (LLMs) involves assessing generated text based on readers' reactions rather than merely its indistinguishability from human-produced content. This paper explores how LLM-generated…

Computational Engineering, Finance, and Science · Computer Science 2024-11-26 Takehiro Takayanagi , Hiroya Takamura , Kiyoshi Izumi , Chung-Chi Chen

We introduce the Generalized Turing Test (GTT), a formal framework for comparing the capabilities of arbitrary agents via indistinguishability. For agents A and B, we define the Turing comparator A $\geq$ B to hold if B, acting as a…

Artificial Intelligence · Computer Science 2026-05-12 Daniel Mitropolsky , Susan S. Hong , Riccardo Neumarker , Emanuele Rimoldi , Tomaso Poggio

As AI-powered image generation improves, a key question is how well human beings can differentiate between "real" and AI-generated or modified images. Using data collected from the online game "Real or Not Quiz.", this study investigates…

Human-Computer Interaction · Computer Science 2025-07-28 Thomas Roca , Anthony Cintron Roman , Jehú Torres Vega , Marcelo Duarte , Pengce Wang , Kevin White , Amit Misra , Juan Lavista Ferres

Matrix Games are a type of unconstrained wargame used by planners to explore scenarios. Players propose actions, and give arguments and counterarguments for their success. An umpire, assisted by dice rolls modified according to the offered…

Computer Science and Game Theory · Computer Science 2024-05-21 Lewis D Griffin , Nicholas Riggs

Algorithmic systems are increasingly deployed to make decisions in many areas of people's lives. The shift from human to algorithmic decision-making has been accompanied by concern about potentially opaque decisions that are not aligned…

We investigate the presence of cognitive biases in three large language models (LLMs): GPT-4o, Gemma 2, and Llama 3.1. The study uses 1,500 experiments across nine established cognitive biases to evaluate the models' responses and…

Artificial Intelligence · Computer Science 2025-09-11 Payam Saeedi , Mahsa Goodarzi , M Abdullah Canbaz

Collaborative decision-making with artificial intelligence (AI) agents presents opportunities and challenges. While human-AI performance often surpasses that of individuals, the impact of such technology on human behavior remains…

Artificial Intelligence · Computer Science 2024-11-18 Marco Matarese , Francesco Rea , Katharina J. Rohlfing , Alessandra Sciutti

ChatGPT has shown its great power in text processing, including its reasoning ability from text reading. However, there has not been any direct comparison between human readers and ChatGPT in reasoning ability related to text reading. This…

Computation and Language · Computer Science 2023-11-20 Tongquan Zhou , Yao Zhang , Siyi Cao , Yulu Li , Tao Wang