English
Related papers

Related papers: Learning Optimal Contracts: How to Exploit Small A…

200 papers

We consider the problem of contextual bandits and imitation learning, where the learner lacks direct knowledge of the executed action's reward. Instead, the learner can actively query an expert at each round to compare two actions and…

Machine Learning · Computer Science 2023-07-25 Ayush Sekhari , Karthik Sridharan , Wen Sun , Runzhe Wu

We analyze a two-period principal-agent model in which the principal faces a budget constraint, and the agent's private costs of performing tasks across the two periods may be correlated. We examine the optimal design of the reward scheme…

Optimization and Control · Mathematics 2025-12-30 Eilon Solan , Avraham Tabbach , Chang Zhao

An agent observes the set of available projects and proposes some, but not necessarily all, of them. A principal chooses one or none from the proposed set. We solve for a mechanism that minimizes the principal's worst-case regret. We…

Theoretical Economics · Economics 2023-09-04 Yingni Guo , Eran Shmaya

In the combinatorial action model of contract design, a principal delegates a complex project to an agent, incentivizing a subset of actions from a ground set of $n$ actions, via a linear contract. Computing the optimal contract is a…

Computer Science and Game Theory · Computer Science 2026-04-17 Elizabeth Baldwin , Paul Duetting , Michal Feldman , Maya Schlesinger

We study the problem of learning the utility functions of no-regret learning agents in a repeated normal-form game. Differing from most prior literature, we introduce a principal with the power to observe the agents playing the game, send…

Computer Science and Game Theory · Computer Science 2026-05-13 Brian Hu Zhang , Tao Lin , Yiling Chen , Tuomas Sandholm

We address the problem of learning in an online setting where the learner repeatedly observes features, selects among a set of actions, and receives reward for the action taken. We provide the first efficient algorithm with an optimal…

Machine Learning · Computer Science 2011-06-17 Miroslav Dudik , Daniel Hsu , Satyen Kale , Nikos Karampatziakis , John Langford , Lev Reyzin , Tong Zhang

We study a new class of contract design problems where a principal delegates the execution of multiple projects to a set of agents. The principal's expected reward from each project is a combinatorial function of the agents working on it.…

Computer Science and Game Theory · Computer Science 2025-06-09 Tal Alon , Matteo Castiglioni , Junjie Chen , Tomer Ezra , Yingkai Li , Inbal Talgam-Cohen

The widespread deployment of Machine Learning systems everywhere raises challenges, such as dealing with interactions or competition between multiple learners. In that goal, we study multi-agent sequential decision-making by considering…

Computer Science and Game Theory · Computer Science 2025-10-28 Antoine Scheid , Etienne Boursier , Alain Durmus , Eric Moulines , Michael I. Jordan

We study a bilevel \emph{max-max} optimization framework for principal-agent contract design, in which a principal chooses incentives to maximize utility while anticipating the agent's best response. This problem, central to moral hazard…

Machine Learning · Computer Science 2025-10-27 Tomer Galanti , Aarya Bookseller , Korok Ray

We study a general class of Principal-Agent problems in continuous time under hidden action. By formulating the model as a coupled stochastic optimal control problem we are able to find a set of necessary conditions characterizing optimal…

Optimization and Control · Mathematics 2014-11-27 Boualem Djehiche , Peter Helgesson

In an online contract selection problem there is a seller which offers a set of contracts to sequentially arriving buyers whose types are drawn from an unknown distribution. If there exists a profitable contract for the buyer in the offered…

Machine Learning · Computer Science 2013-05-16 Cem Tekin , Mingyan Liu

We study online learning settings in which experts act strategically to maximize their influence on the learning algorithm's predictions by potentially misreporting their beliefs about a sequence of binary events. Our goal is twofold.…

Machine Learning · Computer Science 2020-07-02 Rupert Freeman , David M. Pennock , Chara Podimata , Jennifer Wortman Vaughan

We study an online mixed discrete and continuous optimization problem where a decision maker interacts with an unknown environment for a number of $T$ rounds. At each round, the decision maker needs to first jointly choose a discrete and a…

Optimization and Control · Mathematics 2024-08-27 Lintao Ye , Ming Chi , Zhi-Wei Liu , Xiaoling Wang , Vijay Gupta

In the combinatorial-action contract model (D\"utting et al., FOCS'21) a principal delegates the execution of a complex project to an agent, who can choose any subset from a given set of actions. Each set of actions incurs a cost to the…

Computer Science and Game Theory · Computer Science 2025-11-27 Paul Dütting , Michal Feldman , Yoav Gal-Tzur , Aviad Rubinstein

We consider the problem of a learning agent who has to repeatedly play a general sum game against a strategic opponent who acts to maximize their own payoff by optimally responding against the learner's algorithm. The learning agent knows…

Computer Science and Game Theory · Computer Science 2025-02-21 Eshwar Ram Arunachaleswaran , Natalie Collina , Jon Schneider

We study multi-agent contract design, where a principal incentivizes a team of agents to take costly actions that jointly determine the project success via a combinatorial reward function. While prior work largely focuses on unconstrained…

Computer Science and Game Theory · Computer Science 2026-03-10 Michal Feldman , Yoav Gal-Tzur , Tomasz Ponitka , Maya Schlesinger

In this paper, we consider a problem of contract theory in which several Principals hire a common Agent and we study the model in the continuous time setting. We show that optimal contracts should satisfy some equilibrium conditions and we…

Optimization and Control · Mathematics 2018-01-15 Thibaut Mastrolia , Zhenjie Ren

In this paper we consider multi-objective reinforcement learning where the objectives are balanced using preferences. In practice, the preferences are often given in an adversarial manner, e.g., customers can be picky in many applications.…

Machine Learning · Computer Science 2021-10-29 Jingfeng Wu , Vladimir Braverman , Lin F. Yang

Firms increasingly delegate decisions to learning algorithms in platform markets. Standard algorithms perform well when platform policies are stationary, but firms often face ambiguity about whether policies are stationary or adapt…

Theoretical Economics · Economics 2026-02-11 Kyohei Okumura

We present a study on a repeated delegated choice problem, which is the first to consider an online learning variant of Kleinberg and Kleinberg, EC'18. In this model, a principal interacts repeatedly with an agent who possesses an exogenous…

Computer Science and Game Theory · Computer Science 2024-02-15 MohammadTaghi Hajiaghayi , Mohammad Mahdavi , Keivan Rezaei , Suho Shin