English
Related papers

Related papers: Successive Incentives

200 papers

Influential benchmarks incentivize competing model developers to strategically allocate post-training resources toward improvements on the leaderboard, a phenomenon dubbed benchmaxxing or training on the test task. In this work, we initiate…

Computer Science and Game Theory · Computer Science 2026-03-10 Yatong Chen , Guanhua Zhang , Moritz Hardt

We introduce and study incentive equilibria for multi-player meanpayoff games. Incentive equilibria generalise well-studied solution concepts such as Nash equilibria and leader equilibria (also known as Stackelberg equilibria). Recall that…

Computer Science and Game Theory · Computer Science 2015-11-03 Anshul Gupta , M. S. Krishna Deepak , Bharath Kumar Padarthi , Sven Schewe , Ashutosh Trivedi

A simple model for cooperation between "selfish" agents, which play an extended version of the Prisoner's Dilemma(PD) game, in which they use arbitrary payoffs, is presented and studied. A continuous variable, representing the probability…

Condensed Matter · Physics 2009-11-10 H. Fort

In the future, artificial learning agents are likely to become increasingly widespread in our society. They will interact with both other learning agents and humans in a variety of complex settings including social dilemmas. We argue that…

Artificial Intelligence · Computer Science 2022-02-22 Tobias Baumann

Sequential Bayesian experimental design typically assumes that the number of experiments is fixed before data collection begins. In practical campaigns, however, experimentation may need to terminate early because additional measurements…

Methodology · Statistics 2026-05-29 Chen Cheng , Xun Huan

This report investigates the optimal design of event-triggered estimation for first-order linear stochastic systems. The problem is posed as a two-player team problem with a partially nested information pattern. The two players are given by…

Optimization and Control · Mathematics 2012-03-23 Adam Molin , Sandra Hirche

We study a continuous time contracting model in which a principal hires a risk averse agent to manage a project over a finite horizon and provides sequential payments whose timing is endogenously determined. The resulting nonzero-sum…

Theoretical Economics · Economics 2025-12-01 Guillermo Alonso Alvarez , Ibrahim Ekren , Liwei Huang

I study sequential contests where the efforts of earlier players may be disclosed to later players by nature or by design. The model has a range of applications, including rent seeking, R&D, oligopoly, public goods provision, and tragedy of…

Computer Science and Game Theory · Computer Science 2021-02-22 Toomas Hinnosaar

The problem of reward design examines the interaction between a leader and a follower, where the leader aims to shape the follower's behavior to maximize the leader's payoff by modifying the follower's reward function. Current approaches to…

Optimization and Control · Mathematics 2024-06-10 Shuo Wu , Haoxiang Ma , Jie Fu , Shuo Han

This paper studies a class of strongly monotone games involving non-cooperative agents that optimize their own time-varying cost functions. We assume that the agents can observe other agents' historical actions and choose actions that best…

Optimization and Control · Mathematics 2023-09-04 Zifan Wang , Yi Shen , Michael M. Zavlanos , Karl H. Johansson

We obtain global, non-asymptotic convergence guarantees for independent learning algorithms in competitive reinforcement learning settings with two agents (i.e., zero-sum stochastic games). We consider an episodic setting where in each…

Machine Learning · Computer Science 2021-01-13 Constantinos Daskalakis , Dylan J. Foster , Noah Golowich

The latest developments in AI focus on agentic systems where artificial and human agents cooperate to realize global goals. An example is collaborative learning, which aims to train a global model based on data from individual agents. A…

Computer Science and Game Theory · Computer Science 2025-08-20 Björn Filter , Ralf Möller , Özgür Lütfü Özçep

Designing distributed algorithms for multi-agent problems is vital for many emerging application domains, and game-theoretic approaches are emerging as a useful paradigm to design such algorithms. However, much of the emphasis of the…

Systems and Control · Electrical Eng. & Systems 2024-05-03 Rohit Konda , Rahul Chandan , David Grimsman , Jason R. Marden

This paper proposes models of learning process in teams of individuals who collectively execute a sequence of tasks and whose actions are determined by individual skill levels and networks of interpersonal appraisals and influence. The…

Social and Information Networks · Computer Science 2016-10-03 Wenjun Mei , Noah E. Friedkin , Kyle Lewis , Francesco Bullo

This paper develops a framework for the design of scoring rules to optimally incentivize an agent to exert a multi-dimensional effort. This framework is a generalization to strategic agents of the classical knapsack problem (cf. Briest,…

Computer Science and Game Theory · Computer Science 2023-07-03 Jason D. Hartline , Liren Shan , Yingkai Li , Yifan Wu

We introduce the class of pay or play games, which captures scenarios in which each decision maker is faced with a choice between two actions: one with a fixed payoff and an- other with a payoff dependent on others' selected actions. This…

Computer Science and Game Theory · Computer Science 2013-09-27 Sigal Oren , Michael Schapira , Moshe Tennenholtz

We study a problem where a group of agents has to decide how a joint reward should be shared among them. We focus on settings where the share that each agent receives depends on the subjective opinions of its peers concerning that agent's…

Computer Science and Game Theory · Computer Science 2013-05-23 Arthur Carvalho , Kate Larson

This work suggests modifications to a previously introduced class of heterogeneous agent models that allow for the inclusion of different types of agent motivations and behaviours in a unified way. The agents operate within a highly…

Trading and Market Microstructure · Quantitative Finance 2009-11-13 H. Lamba , T. Seaman

Designers of AI agents often iterate on the reward function in a trial-and-error process until they get the desired behavior, but this only guarantees good behavior in the training environment. We propose structuring this process as a…

Machine Learning · Computer Science 2023-10-17 Sören Mindermann , Rohin Shah , Adam Gleave , Dylan Hadfield-Menell

Sequence models are a critical component of modern NLP systems, but their predictions are difficult to explain. We consider model explanations though rationales, subsets of context that can explain individual model predictions. We find…

Computation and Language · Computer Science 2021-11-19 Keyon Vafa , Yuntian Deng , David M. Blei , Alexander M. Rush