English
Related papers

Related papers: Measuring Cognitive Abilities in the Wild: Validat…

200 papers

Learning requires the traversal of inherently distinct cognitive states to produce behavioral adaptation. Yet, tools to explicitly measure these states with non-invasive imaging -- and to assess their dynamics during learning -- remain…

Accurately estimating human skill levels is crucial for designing effective human-AI interactions so that AI can provide appropriate challenges or guidance. In games where AI players have beaten top human professionals, strength estimation…

Machine Learning · Computer Science 2025-05-02 Kyota Kuboki , Tatsuyoshi Ogawa , Chu-Hsuan Hsueh , Shi-Jim Yen , Kokolo Ikeda

Outcome-driven studies designed to evaluate potential effects of games and apps designed to promote healthy eating and exercising remain limited either targeting design or usability factors while omitting out health-based outcomes…

Human-Computer Interaction · Computer Science 2021-06-28 S. Durga , S. Hallinan , M. Seif El-Nasr , M. Shiyko , C. Sceppa

Previous work has shown that it is possible to train neuronal cultures on Multi-Electrode Arrays (MEAs), to recognize very simple patterns. However, this work was mainly focused to demonstrate that it is possible to induce plasticity in…

Computer Vision and Pattern Recognition · Computer Science 2021-01-25 Gabriele Lagani , Raffaele Mazziotti , Fabrizio Falchi , Claudio Gennaro , Guido Marco Cicchini , Tommaso Pizzorusso , Federico Cremisi , Giuseppe Amato

Cognitive effort, defined as the relationship between cognitive load and task performance, provides insight into how individuals allocate mental resources during demanding tasks. This construct is particularly important in high-stakes…

Human-Computer Interaction · Computer Science 2026-04-13 Shayla Sharmin , Mohammad Fahim Abrar , Gael Lucero-Palacios , Aditya Raikwar , Roghayeh Leila Barmaki

Categories such as animal or furniture are acquired at an early age and play an important role in processing, organizing, and communicating world knowledge. Categories exist across cultures: they allow to efficiently represent the…

Computation and Language · Computer Science 2019-02-26 Lea Frermann , Mirella Lapata

Conversational Spoken Language Models (SLMs) are emerging as a promising paradigm for real-time speech interaction. However, their capacity of temporal dynamics, including the ability to manage timing, tempo and simultaneous speaking,…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-04 Kai-Wei Chang , En-Pei Hu , Chun-Yi Kuan , Wenze Ren , Wei-Chih Chen , Guan-Ting Lin , Yu Tsao , Shao-Hua Sun , Hung-yi Lee , James Glass

Recent benchmark studies have claimed that AI has approached or even surpassed human-level performances on various cognitive tasks. However, this position paper argues that current AI evaluation paradigms are insufficient for assessing…

Psychlab is a simulated psychology laboratory inside the first-person 3D game world of DeepMind Lab (Beattie et al. 2016). Psychlab enables implementations of classical laboratory psychological experiments so that they work with both human…

The analysis of the adaptive behaviour of many different kinds of systems such as humans, animals and machines, requires more general ways of assessing their cognitive abilities. This need is strengthened by increasingly more tasks being…

Artificial Intelligence · Computer Science 2013-05-10 David L. Dowe , Jose Hernandez-Orallo

Self-assessment is a key aspect of reliable intelligence, yet evaluations of large language models (LLMs) focus mainly on task accuracy. We adapted the 10-item General Self-Efficacy Scale (GSES) to elicit simulated self-assessments from ten…

Artificial Intelligence · Computer Science 2025-11-27 Daniel I Jackson , Emma L Jensen , Syed-Amad Hussain , Emre Sezgin

There has been a growing interest in using AI to model human behavior, particularly in domains where humans interact with this technology. While most existing work models human behavior at an aggregate level, our goal is to model behavior…

Machine Learning · Computer Science 2025-02-24 Nabil Omi , Lucas Caccia , Anurag Sarkar , Jordan T. Ash , Siddhartha Sen

Psychological measurement is essential for mental health, self-understanding, and personal development. Traditional methods, such as self-report scales and psychologist interviews, often face challenges with engagement and accessibility.…

Computation and Language · Computer Science 2024-08-30 Qisen Yang , Zekun Wang , Honghui Chen , Shenzhi Wang , Yifan Pu , Xin Gao , Wenhao Huang , Shiji Song , Gao Huang

Traditional cognitive bias measurement tools are limited by narrow bias coverage, low ecological validity, and reliance on abstract self reports, constraining scenario based and human AI comparisons. We introduce the context based Cognitive…

Human-Computer Interaction · Computer Science 2026-04-28 Chengrui Zhou

Serious games have proven to be effective tools for screening cognitive impairments and supporting diagnosis in patients with neurodegenerative diseases like Alzheimer's and Parkinson's. They also offer cognitive training benefits.…

Formal Languages and Automata Theory · Computer Science 2026-02-04 Elisabetta De Maria , Christopher Leturc

Recent advances in deep reinforcement learning (RL) have demonstrated complex decision-making capabilities in simulation environments such as Arcade Learning Environment, MuJoCo, and ViZDoom. However, they are hardly extensible to more…

Machine Learning · Computer Science 2022-10-18 Xi Chen , Tianyu Shi , Qingpeng Zhao , Yuchen Sun , Yunfei Gao , Xiangjun Wang

Intelligence is a crucial trait for species to find solutions within a limited number of trial-and-error attempts. Building on this idea, we introduce Survival Game as a framework to evaluate intelligence based on the number of failed…

Artificial Intelligence · Computer Science 2025-03-06 Jingtao Zhan , Jiahao Zhao , Jiayu Li , Yiqun Liu , Bo Zhang , Qingyao Ai , Jiaxin Mao , Hongning Wang , Min Zhang , Shaoping Ma

Recording the dynamics of unscripted human interactions in the wild is challenging due to the delicate trade-offs between several factors: participant privacy, ecological validity, data fidelity, and logistical overheads. To address these,…

Multimedia · Computer Science 2022-10-11 Chirag Raman , Jose Vargas-Quiros , Stephanie Tan , Ashraful Islam , Ekin Gedik , Hayley Hung

Comparing AI models to "human level" is often misleading when benchmark scores are incommensurate or human baselines are drawn from a narrow population. To address this, we propose a framework that calibrates items against the 'world…

Playing games is inherently human, and a lot of games are created to challenge different human characteristics. However, these tasks are often left out when evaluating the human-like nature of artificial models. The objective of this work…

Computer Vision and Pattern Recognition · Computer Science 2025-09-04 Nuria Alabau-Bosque , Jorge Vila-Tomás , Paula Daudén-Oliver , Pablo Hernández-Cámara , Jose Manuel Jaén-Lorites , Valero Laparra , Jesús Malo