English
Related papers

Related papers: The Burden of Interactive Alignment with Inconsist…

200 papers

We study partially observable assistance games (POAGs), a model of the human-AI value alignment problem which allows the human and the AI assistant to have partial observations. Motivated by concerns of AI deception, we study a…

Artificial Intelligence · Computer Science 2025-08-12 Scott Emmons , Caspar Oesterheld , Vincent Conitzer , Stuart Russell

High-stakes applications rely on combining Artificial Intelligence (AI) and humans for responsive and reliable decision making. For example, content moderation in social media platforms often employs an AI-human pipeline to promptly remove…

Machine Learning · Computer Science 2025-08-14 Thodoris Lykouris , Wentao Weng

As reliance on AI systems for decision-making grows, it becomes critical to ensure that human users can appropriately balance trust in AI suggestions with their own judgment, especially in high-stakes domains like healthcare. However, human…

Human-Computer Interaction · Computer Science 2025-01-29 Zichen Chen , Yunhao Luo , Misha Sra

Optimizing recommender systems based on user interaction data is mainly seen as a problem of dealing with selection bias, where most existing work assumes that interactions from different users are independent. However, it has been shown…

Information Retrieval · Computer Science 2022-07-04 Norman Knyazev , Harrie Oosterhuis

Alignment faking is a form of strategic deception in AI in which models selectively comply with training objectives when they infer that they are in training, while preserving different behavior outside training. The phenomenon was first…

As agents based on large language models are increasingly deployed to long-horizon tasks, maintaining their alignment with stakeholder preferences becomes critical. Effective alignment in such settings requires reward models that are…

Artificial Intelligence · Computer Science 2025-12-09 Charlie Masters , Marta Grześkiewicz , Stefano V. Albrecht

Human-AI interactions are increasingly part of everyday life, yet the interpersonal dynamics that unfold during such exchanges remain underexplored. This study investigates how emotional alignment, semantic exploration, and linguistic…

Human-Computer Interaction · Computer Science 2025-12-22 Halfdan Nordahl Fundal , Johannes Eide Rambøll , Karsten Olsen

Aligning large language models with humans is challenging due to the inherently multifaceted nature of preference feedback. While existing approaches typically frame this as a multi-objective optimization problem, they often overlook how…

Computation and Language · Computer Science 2025-06-03 Mohamad Chehade , Soumya Suvra Ghosal , Souradip Chakraborty , Avinash Reddy , Dinesh Manocha , Hao Zhu , Amrit Singh Bedi

This paper provides a simple theoretical framework to evaluate the effect of key parameters of ranking algorithms, namely popularity and personalization parameters, on measures of platform engagement, misinformation and polarization. The…

Social and Information Networks · Computer Science 2022-10-06 Fabrizio Germano , Vicenç Gómez , Francesco Sobbrio

Large language models improve with scale, yet feedback-based alignment still exhibits systematic deviations from intended behavior. Motivated by bounded rationality in economics and cognitive science, we view judgment as resource-limited…

Machine Learning · Computer Science 2025-09-22 Wenjun Cao

As predictive models are deployed into the real world, they must increasingly contend with strategic behavior. A growing body of work on strategic classification treats this problem as a Stackelberg game: the decision-maker "leads" in the…

Machine Learning · Computer Science 2022-02-01 Tijana Zrnic , Eric Mazumdar , S. Shankar Sastry , Michael I. Jordan

This paper considers a non-cooperative game in which competing users sharing a frequency-selective interference channel selfishly optimize their power allocation in order to improve their achievable rates. Previously, it was shown that a…

Computer Science and Game Theory · Computer Science 2008-11-04 Yi Su , Mihaela van der Schaar

Recent advances in large language models have highlighted their potential for personalized recommendation, where accurately capturing user preferences remains a key challenge. Leveraging their strong reasoning and generalization…

As algorithms increasingly mediate competitive decision-making, their influence extends beyond individual outcomes to shaping strategic market dynamics. In two preregistered experiments, we examined how algorithmic advice affects human…

Human-Computer Interaction · Computer Science 2025-11-13 Tobias R. Rebholz , Maxwell Uphoff , Christian H. R. Bernges , Florian Scholten

AI alignment is often framed as the task of ensuring that an AI system follows a set of stated principles or human preferences, but general principles rarely determine their own application in concrete cases. When principles conflict, when…

Artificial Intelligence · Computer Science 2026-04-14 Behrooz Razeghi

In this paper we study the problem of information sharing among rational self-interested agents as a dynamic game of asymmetric information. We assume that the agents imperfectly observe a Markov chain and they are called to decide whether…

Computer Science and Game Theory · Computer Science 2021-03-30 Konstantinos Ntemos , George Pikramenos , Nicholas Kalouptsidis

Large Language Models (LLMs) have achieved remarkable progress in reasoning, yet sometimes produce responses that are suboptimal for users in tasks such as writing, information seeking, or providing practical guidance. Conventional…

Artificial Intelligence · Computer Science 2025-11-04 Siqi Zhu , David Zhang , Pedro Cisneros-Velarde , Jiaxuan You

Millions of people now turn to artificial intelligence (AI) systems for personal advice, guidance, and support. Such systems can be sycophantic, frequently affirming users' views and beliefs. Across five preregistered studies (N = 3,075…

Human-Computer Interaction · Computer Science 2026-05-13 Lujain Ibrahim , Franziska Sofia Hafner , Myra Cheng , Cinoo Lee , Rebecca Anselmetti , Robb Willer , Luc Rocher , Diyi Yang

We consider a setting where $n$ buyers, with combinatorial preferences over $m$ items, and a seller, running a priority-based allocation mechanism, repeatedly interact. Our goal, from observing limited information about the results of these…

Computer Science and Game Theory · Computer Science 2014-08-29 Avrim Blum , Yishay Mansour , Jamie Morgenstern

In a multi-follower Bayesian Stackelberg game, a leader plays a mixed strategy over $L$ actions to which $n\ge 1$ followers, each having one of $K$ possible private types, best respond. The leader's optimal strategy depends on the…

Computer Science and Game Theory · Computer Science 2026-03-03 Gerson Personnat , Tao Lin , Safwan Hossain , David C. Parkes