English
Related papers

Related papers: Is General-Purpose AI Reasoning Sensitive to Data-…

200 papers

Political biases in Large Language Model (LLM)-based artificial intelligence (AI) systems, such as OpenAI's ChatGPT or Google's Gemini, have been previously reported. While several prior studies have attempted to quantify these biases using…

Computers and Society · Computer Science 2025-03-17 David Rozado

This document presents a preliminary compilation of general-purpose AI (GPAI) evaluation practices that may promote internal validity, external validity and reproducibility. It includes suggestions for human uplift studies and benchmark…

Computers and Society · Computer Science 2025-08-20 Patricia Paskov , Michael J. Byun , Kevin Wei , Toby Webster

The significant advancements in applying Artificial Intelligence (AI) to healthcare decision-making, medical diagnosis, and other domains have simultaneously raised concerns about the fairness and bias of AI systems. This is particularly…

Computers and Society · Computer Science 2024-01-30 Emilio Ferrara

Recent advances in large language models (LLMs) have enabled automatic generation of chain-of-thought (CoT) reasoning, leading to strong performance on tasks such as math and code. However, when reasoning steps reflect social stereotypes…

Computation and Language · Computer Science 2025-09-23 Xuyang Wu , Jinming Nian , Ting-Ruen Wei , Zhiqiang Tao , Hsin-Tai Wu , Yi Fang

AI safety is a rapidly growing area of research that seeks to prevent the harm and misuse of frontier AI technology, particularly with respect to generative AI (GenAI) tools that are capable of creating realistic and high-quality content…

Artificial Intelligence · Computer Science 2025-02-19 Pin-Yu Chen

This research explores how human-defined goals influence the behavior of Large Language Models (LLMs) through purpose-conditioned cognition. Using financial prediction tasks, we show that revealing the downstream use (e.g., predicting stock…

General Finance · Quantitative Finance 2026-05-07 Sean Cao , Wei Jiang , Hui Xu

Several strands of research have aimed to bridge the gap between artificial intelligence (AI) and human decision-makers in AI-assisted decision-making, where humans are the consumers of AI model predictions and the ultimate decision-makers…

Human-Computer Interaction · Computer Science 2022-04-06 Charvi Rastogi , Yunfeng Zhang , Dennis Wei , Kush R. Varshney , Amit Dhurandhar , Richard Tomsett

AI is redefining how humans interact with technology, leading to a synergetic collaboration between the two. Nevertheless, the effects of human cognition on this collaboration remain unclear. This study investigates the implications of two…

Human-Computer Interaction · Computer Science 2024-11-12 Samuel Aleksander Sánchez Olszewski

Do generative AI models, particularly large language models (LLMs), exhibit systematic behavioral biases in economic and financial decisions? If so, how can these biases be mitigated? Drawing on the cognitive psychology and experimental…

General Economics · Economics 2026-02-11 Pietro Bini , Lin William Cong , Xing Huang , Lawrence J. Jin

Human decision-making is strongly influenced by cognitive biases, particularly under conditions of uncertainty and risk. While prior work has examined bias in single-step decisions with immediate outcomes and in human interaction with a…

Human-Computer Interaction · Computer Science 2026-03-25 Teerthaa Parakh , Karen M. Feigh

Large language models (LLMs) are able to engage in natural-sounding conversations with humans, showcasing unprecedented capabilities for information retrieval and automated decision support. They have disrupted human-technology interaction…

Computation and Language · Computer Science 2025-02-18 Wolfgang Messner , Tatum Greene , Josephine Matalone

Can AI be cognitively biased in automated information judgment tasks? Despite recent progresses in measuring and mitigating social and algorithmic biases in AI and large language models (LLMs), it is not clear to what extent LLMs behave…

Information Retrieval · Computer Science 2024-11-26 Jiqun Liu , Jiangen He

Although general-purpose artificial intelligence (GPAI) is widely expected to accelerate scientific discovery, its practical limits in biomedicine remain unclear. We assess this potential by developing a framework of GPAI capabilities…

Computers and Society · Computer Science 2025-08-26 Konstantin Hebenstreit , Constantin Convalexius , Stephan Reichl , Stefan Huber , Christoph Bock , Matthias Samwald

Biases in Artificial Intelligence (AI) or Machine Learning (ML) systems due to skewed datasets problematise the application of prediction models in practice. Representation bias is a prevalent form of bias found in the majority of datasets.…

Human-Computer Interaction · Computer Science 2024-07-16 Aditya Bhattacharya , Simone Stumpf , Katrien Verbert

Teaching unbiased decision-making is crucial for addressing biased decision-making in daily life. Although both raising awareness of personal biases and providing guidance on unbiased decision-making are essential, the latter topics remains…

Human-Computer Interaction · Computer Science 2024-04-09 Mingzhe Yang , Hiromi Arai , Naomi Yamashita , Yukino Baba

The rapid proliferation and deployment of General-Purpose AI (GPAI) models, including large language models (LLMs), present unprecedented challenges for AI supervisory entities. We hypothesize that these entities will need to navigate an…

Artificial Intelligence · Computer Science 2025-06-12 Manuel Cebrian , Emilia Gomez , David Fernandez Llorca

Artificial Intelligence (AI) systems are attracting increasing interest in the medical domain due to their ability to learn complicated tasks that require human intelligence and expert knowledge. AI systems that utilize high-performance…

Computation and Language · Computer Science 2021-08-30 Milad Moradi , Kathrin Blagec , Matthias Samwald

We are witnessing the emergence of an AI economy and society where AI technologies are increasingly impacting health care, business, transportation and many aspects of everyday life. Many successes have been reported where AI systems even…

Machine Learning · Computer Science 2022-12-27 D. Petkovic

Comprehensive and accurate evaluation of general-purpose AI systems such as large language models allows for effective mitigation of their risks and deepened understanding of their capabilities. Current evaluation methodology, mostly based…

Artificial Intelligence · Computer Science 2024-01-01 Xiting Wang , Liming Jiang , Jose Hernandez-Orallo , David Stillwell , Luning Sun , Fang Luo , Xing Xie

Generative AI systems have entered everyday academic, professional, and personal life with remarkable speed, yet most users encounter them as mysterious artifacts rather than intelligible systems. This chapter discusses large language…

Computers and Society · Computer Science 2026-04-21 John T. Behrens