English
Related papers

Related papers: The Limits of Limited Commitment

200 papers

This paper investigates a distributed goal assignment problem in leader-following formation control of second-order multi-agent systems. It is assumed that each agent can communicate with nearby agents within the communication range and the…

Systems and Control · Electrical Eng. & Systems 2022-03-03 Yun Ho Choi , Doik Kim

Effective game-theoretic modeling of defender-attacker behavior is becoming increasingly important. In many domains, the defender functions not only as a player but also the designer of the game's payoff structure. We study Stackelberg…

Computer Science and Game Theory · Computer Science 2018-05-23 Zheyuan Ryan Shi , Ziye Tang , Long Tran-Thanh , Rohit Singh , Fei Fang

When multiple informative equilibria are possible in a general cheap talk game, how much information can a principal guarantee herself? To answer this question, I define the notion of worst-case implementation-implementation via the worst…

Theoretical Economics · Economics 2026-02-17 Andrei Iakovlev

Empirical coordination offers a way to understand how agents can coordinate actions under communication constraints. This paper investigates the finite blocklength regime of this problem, where the encoder and decoder aim to produce a…

Information Theory · Computer Science 2026-05-13 Olivier Massicot , Giulia Cervia , Maël Le Treust

We analyze a canonical extension of the Stackelberg duopoly to a sequential framework, where each firm strategically anticipates the reactions of all subsequent players. In a triopoly (three-firm) settings, we obtain existence and…

Functional Analysis · Mathematics 2026-04-30 Anton Badev , Martin Pavlov , Boyan Zlatanov

We study how delegating pricing to large language models (LLMs) can facilitate collusion in a duopoly when both sellers rely on the same pre-trained model. The LLM is characterized by (i) a propensity parameter capturing its internal bias…

Theoretical Economics · Economics 2026-03-24 Shengyu Cao , Ming Hu

This thesis considers sequential decision problems, where the loss/reward incurred by selecting an action may not be inferred from observed feedback. A major part of this thesis focuses on the unsupervised sequential selection problem,…

Machine Learning · Computer Science 2023-01-30 Arun Verma

We study sequential decision making in environments where rewards are only partially observed, but can be modeled as a function of observed contexts and the chosen action by the decision maker. This setting, known as contextual bandits,…

Methodology · Statistics 2015-03-11 Miroslav Dudík , Dumitru Erhan , John Langford , Lihong Li

This paper studies a duopoly investment model with uncertainty. There are two alternative irreversible investments. The first firm to invest gets a monopoly benefit for a specified period of time. The second firm to invest gets information…

Optimization and Control · Mathematics 2019-03-01 Kristina Rognlien Dahl , Espen Stokkereit

We consider a team-production environment where all participants are motivated by career concerns, and where a team's joint productive outcome may have different reputational implications for different team members. In this context, we…

Theoretical Economics · Economics 2023-05-08 Paula Onuchic , João Ramos

The theoretical study of social learning typically assumes that each agent's action affects only her own payoff. In this paper, I present a model in which agents' actions directly affect the payoffs of other agents. On a discrete time line,…

Social and Information Networks · Computer Science 2015-11-02 Yangbo Song

Risk measures are commonly used to capture the risk preferences of decision-makers (DMs). The decisions of DMs can be nudged or manipulated when their risk preferences are influenced by factors such as the availability of information about…

Optimization and Control · Mathematics 2023-11-29 Shutian Liu , Quanyan Zhu

The problem of statistical learning is to construct an accurate predictor of a random variable as a function of a correlated random variable on the basis of an i.i.d. training sample from their joint distribution. Allowable predictors are…

Information Theory · Computer Science 2009-04-30 Maxim Raginsky

We study which outcomes are implementable by disclosing coarse statistics of a data-generating process rather than its full distribution. Players observe data whose joint distribution is only partially known: they know the expectations of…

Theoretical Economics · Economics 2026-05-11 Francesco Giordano

The work studies the problem of decentralized constrained POMDPs in a team-setting where multiple nonstrategic agents have asymmetric information. Using an extension of Sion's Minimax theorem for functions with positive infinity and results…

Optimization and Control · Mathematics 2025-04-29 Nouman Khan , Vijay Subramanian

Formulating a real-world problem under the Reinforcement Learning framework involves non-trivial design choices, such as selecting a discount factor for the learning objective (discounted cumulative rewards), which articulates the planning…

Artificial Intelligence · Computer Science 2025-02-19 Randy Lefebvre , Audrey Durand

Individual components such as cells, particles, or agents within a larger system often require detailed understanding of their relative position to act accordingly, enabling the system as a whole to function in an organised and efficient…

Statistical Mechanics · Physics 2025-02-28 Jonas Berx , Prashant Singh , Karel Proesmans

We present a two-level model of organizational training and agent production. Managers decide whether or not to train based on both the costs of training compared to the benefits and on their expectations and observations of the number of…

adap-org · Physics 2008-02-03 Natalie S. Glance , Tad Hogg , Bernardo A. Huberman

Humans and animals have the ability to reason and make predictions about different courses of action at many time scales. In reinforcement learning, option models (Sutton, Precup \& Singh, 1999; Precup, 2000) provide the framework for this…

Machine Learning · Computer Science 2021-08-09 Khimya Khetarpal , Zafarali Ahmed , Gheorghe Comanici , Doina Precup

In experimental applications of bounded-reasoning models, behavior is often summarized by distributions of "levels". We argue that such summaries conflate two conceptually distinct dimensions: a player's type, capturing beliefs about what…

Theoretical Economics · Economics 2026-04-15 Shuige Liu , Gabriel Ziegler