English
Related papers

Related papers: The Burden of Interactive Alignment with Inconsist…

200 papers

Existing work on the alignment problem has focused mainly on (1) qualitative descriptions of the alignment problem; (2) attempting to align AI actions with human interests by focusing on value specification and learning; and/or (3) focusing…

Multiagent Systems · Computer Science 2025-06-03 Aidan Kierans , Avijit Ghosh , Hananel Hazan , Shiri Dori-Hacohen

Achieving human-AI alignment in complex multi-agent games is crucial for creating trustworthy AI agents that enhance gameplay. We propose a method to evaluate this alignment using an interpretable task-sets framework, focusing on high-level…

Artificial Intelligence · Computer Science 2024-06-21 Sugandha Sharma , Guy Davidson , Khimya Khetarpal , Anssi Kanervisto , Udit Arora , Katja Hofmann , Ida Momennejad

A long-standing vision of computing is the personal AI system: one that understands us well enough to address our underlying needs. Today's AI focuses on what users do, ignoring why they might be doing such things in the first place. As a…

Human-Computer Interaction · Computer Science 2026-04-10 Dora Zhao , Michelle S. Lam , Diyi Yang , Michael S. Bernstein

Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regularization with respect to the base policy. This practice is suboptimal for maximizing user's utility…

Machine Learning · Computer Science 2026-02-04 Haichuan Wang , Tao Lin , Lingkai Kong , Ce Li , Hezi Jiang , Milind Tambe

Stackelberg games have been widely used to model interactive decision-making problems in a variety of domains such as energy systems, transportation, cybersecurity, and human-robot interaction. However, existing algorithms for solving…

Optimization and Control · Mathematics 2023-03-14 Yansong Li , Shuo Han

We study Stackelberg equilibria in finitely repeated games, where the leader commits to a strategy that picks actions in each round and can be adaptive to the history of play (i.e. they commit to an algorithm). In particular, we study…

Computer Science and Game Theory · Computer Science 2024-03-08 Natalie Collina , Eshwar Ram Arunachaleswaran , Michael Kearns

Social media feed algorithms infer user preferences from their past behaviors. Yet what drives engagement often diverges from what users value. We examine this gap between stated preferences (what users say they prefer) and revealed…

Human-Computer Interaction · Computer Science 2026-04-14 Do Won Kim , Cody Buntain , Giovanni Luca Ciampaglia

When deployed in the world, a learning agent such as a recommender system or a chatbot often repeatedly interacts with another learning agent (such as a user) over time. In many such two-agent systems, each agent learns separately and the…

Machine Learning · Computer Science 2024-06-24 Kate Donahue , Nicole Immorlica , Meena Jagadeesan , Brendan Lucier , Aleksandrs Slivkins

We formalize AI alignment as a multi-objective optimization problem called $\langle M,N,\varepsilon,\delta\rangle$-agreement, in which a set of $N$ agents (including humans) must reach approximate ($\varepsilon$) agreement across $M$…

Artificial Intelligence · Computer Science 2025-11-20 Aran Nayebi

Alignment is a social phenomenon wherein individuals share a common goal or perspective. Mirroring, or mimicking the behaviors and opinions of another individual, is one mechanism by which individuals can become aligned. Large scale…

Multiagent Systems · Computer Science 2025-02-18 Harvey McGuinness , Tianyu Wang , Carey E. Priebe , Hayden Helm

Digital platforms such as social media and e-commerce websites adopt Recommender Systems to provide value to the user. However, the social consequences deriving from their adoption are still unclear. Many scholars argue that recommenders…

Information Retrieval · Computer Science 2024-09-26 Erica Coppolillo , Simone Mungari , Ettore Ritacco , Francesco Fabbri , Marco Minici , Francesco Bonchi , Giuseppe Manco

This paper studies a Stackelberg game wherein a sender (leader) attempts to shape the information of a less informed receiver (follower) who in turn takes an action that determines the payoff for both players. The sender chooses signals to…

Computer Science and Game Theory · Computer Science 2022-10-07 Reema Deori , Ankur A. Kulkarni

Shared control allows the human driver to collaborate with an assistive driving system while retaining the ability to make decisions and take control if necessary. However, human-vehicle teaming and planning are challenging due to…

Robotics · Computer Science 2024-03-19 Yuhan Zhao , Quanyan Zhu

When learning in strategic environments, a key question is whether agents can overcome uncertainty about their preferences to achieve outcomes they could have achieved absent any uncertainty. Can they do this solely through interactions…

Computer Science and Game Theory · Computer Science 2024-11-21 Nivasini Ananthakrishnan , Nika Haghtalab , Chara Podimata , Kunhe Yang

In the age of information abundance, attention is a coveted resource. Social media platforms vigorously compete for users' engagement, influencing the evolution of their opinions on a variety of topics. With recommendation algorithms often…

Physics and Society · Physics 2023-10-30 Andrea Somazzi , Giuseppe Maria Ferro , Diego Garlaschelli , Simon Asher Levin

Many online platforms predominantly rank items by predicted user engagement. We believe that there is much unrealized potential in including non-engagement signals, which can improve outcomes both for platforms and for society as a whole.…

Social and Information Networks · Computer Science 2024-02-13 Tom Cunningham , Sana Pandey , Leif Sigerson , Jonathan Stray , Jeff Allen , Bonnie Barrilleaux , Ravi Iyer , Smitha Milli , Mohit Kothari , Behnam Rezaei

In self-consuming generative models that train on their own outputs, alignment with user preferences becomes a recursive rather than one-time process. We provide the first formal foundation for analyzing the long-term effects of such…

Machine Learning · Computer Science 2025-11-18 Ali Falahati , Mohammad Mohammadi Amiri , Kate Larson , Lukasz Golab

This paper considers two investors who perform mean-variance portfolio selection with asymmetric information: one knows the true stock dynamics, while the other has to infer the true dynamics from observed stock evolution. Their portfolio…

Mathematical Finance · Quantitative Finance 2025-09-05 Yu-Jui Huang , Shihao Zhu

Recommendation algorithms play a pivotal role in shaping our media choices, which makes it crucial to comprehend their long-term impact on user behavior. These algorithms are often linked to two critical outcomes: homogenization, wherein…

Computers and Society · Computer Science 2024-03-11 Md Sanzeed Anwar , Grant Schoenebeck , Paramveer S. Dhillon

Many human-facing algorithms -- including those that power recommender systems or hiring decision tools -- are trained on data provided by their users. The developers of these algorithms commonly adopt the assumption that the data…

Computers and Society · Computer Science 2024-01-01 Sarah H. Cen , Andrew Ilyas , Aleksander Madry