中文
相关论文

相关论文: Optimal Incentive Contract with Endogenous Monitor…

200 篇论文

We consider the problem of revenue-optimal dynamic mechanism design in settings where agents' types evolve over time as a function of their (both public and private) experience with items that are auctioned repeatedly over an infinite…

计算机科学与博弈论 · 计算机科学 2010-10-18 Sham M. Kakade , Ilan Lobel , Hamid Nazerzadeh

We introduce a stochastic principal-agent model. A principal and an agent interact in a stochastic environment, each privy to observations about the state not available to the other. The principal has the power of commitment, both to elicit…

计算机科学与博弈论 · 计算机科学 2024-09-13 Jiarui Gan , Rupak Majumdar , Debmalya Mandal , Goran Radanovic

We design an optimal contract between a demand response aggregator (DRA) and a customer for incentive-based demand response. We consider a setting in which the customer is asked to reduce her consumption by the DRA and she is compensated…

最优化与控制 · 数学 2024-10-30 Donya G. Dobakhshari , Vijay Gupta

Recent years are seeing an increasing need for on-line monitoring of teams of cooperating agents, e.g., for visualization, or performance tracking. However, in monitoring deployed teams, we often cannot rely on the agents to always…

人工智能 · 计算机科学 2011-06-10 G. A. Kaminka , D. V. Pynadath , M. Tambe

In many scenarios, a principal dynamically interacts with an agent and offers a sequence of incentives to align the agent's behavior with a desired objective. This paper focuses on the problem of synthesizing an incentive sequence that,…

最优化与控制 · 数学 2020-07-20 Yagiz Savas , Vijay Gupta , Ufuk Topcu

Autonomous agents optimize the reward function we give them. What they don't know is how hard it is for us to design a reward function that actually captures what we want. When designing the reward, we might think of some specific training…

人工智能 · 计算机科学 2020-10-08 Dylan Hadfield-Menell , Smitha Milli , Pieter Abbeel , Stuart Russell , Anca Dragan

We propose an incentive scheme based on intervention to sustain cooperation among self-interested users. In the proposed scheme, an intervention device collects imperfect signals about the actions of the users for a test period, and then…

计算机科学与博弈论 · 计算机科学 2010-12-09 Jaeok Park , Mihaela van der Schaar

Dynamic contracts with multiple agents is a classical decentralized decision-making problem with asymmetric information. In this paper, we extend the single-agent dynamic incentive contract model in continuous-time to a multi-agent scheme…

计量经济学 · 经济学 2017-10-10 Qi Luo , Romesh Saigal

This paper studies the design of optimal proper scoring rules when the principal has partial knowledge of an agent's signal distribution. Recent work characterizes the proper scoring rules that maximize the increase of an agent's payoff…

计算机科学与博弈论 · 计算机科学 2024-10-15 Yiling Chen , Fang-Yi Yu

A principal selects a team of agents for collaborating on a joint project. The principal aims to design a revenue-optimal contract that incentivize the team of agents to exert costly effort while satisfying fairness constraints. We show…

计算机科学与博弈论 · 计算机科学 2025-12-23 Matteo Castiglioni , Junjie Chen , Yingkai Li

The world is full of systems of distributed agents, collaborating and competing in complex ways: firms and workers specialise within economies, neurons adapt their tuning across brain circuits, and species compete and coexist within…

神经与进化计算 · 计算机科学 2026-04-10 Guillhem Artis , Danyal Akarca , Jascha Achterberg

We study a dynamic contracting problem with multiple agents and limited commitment. A principal seeks to screen efficient agents using one-period contracts, but is tempted to revise contract terms upon knowing an agent's type. Alterations…

理论经济学 · 经济学 2025-03-24 Mehmet Ekmekci , Lucas Maestri , Dong Wei

A robust adaptive model predictive control (MPC) algorithm is presented for linear, time invariant systems with unknown dynamics and subject to bounded measurement noise. The system is characterized by an impulse response model, which is…

系统与控制 · 电气工程与系统科学 2019-11-21 Anilkumar Parsi , Andrea Iannelli , Mingzhou Yin , Mohammad Khosravi , Roy S. Smith

Agents in dynamic multi-agent environments must monitor their peers to execute individual and group plans. A key open question is how much monitoring of other agents' states is required to be effective: The Monitoring Selectivity Problem.…

多智能体系统 · 计算机科学 2011-06-02 G. A. Kaminka , M. Tambe

We introduce a new model of combinatorial contracts in which a principal delegates the execution of a costly task to an agent. To complete the task, the agent can take any subset of a given set of unobservable actions, each of which has an…

计算机科学与博弈论 · 计算机科学 2025-09-03 Paul Duetting , Tomer Ezra , Michal Feldman , Thomas Kesselheim

We study multi-agent contract design, where a principal incentivizes a team of agents to take costly actions that jointly determine the project success via a combinatorial reward function. While prior work largely focuses on unconstrained…

计算机科学与博弈论 · 计算机科学 2026-03-10 Michal Feldman , Yoav Gal-Tzur , Tomasz Ponitka , Maya Schlesinger

Learning how to learn efficiently is a fundamental challenge for biological agents and a growing concern for artificial ones. To learn effectively, an agent must regulate its learning speed, balancing the benefits of rapid improvement…

机器学习 · 计算机科学 2026-01-13 Valentina Njaradi , Rodrigo Carrasco-Davis , Peter E. Latham , Andrew Saxe

Combined prosocial incentives, integrating reward for cooperators and punishment for defectors, are effective tools to promote cooperation among competing agents in population games. Existing research concentrated on how to adjust reward or…

最优化与控制 · 数学 2023-12-06 Shengxian Wang , Ming Cao , Xiaojie Chen

We initiate the study of computing (near-)optimal contracts in succinctly representable principal-agent settings. Here optimality means maximizing the principal's expected payoff over all incentive-compatible contracts---known in economics…

数据结构与算法 · 计算机科学 2020-02-28 Paul Duetting , Tim Roughgarden , Inbal Talgam-Cohen

Selective labels are a common feature of consequential decision-making applications, referring to the lack of observed outcomes under one of the possible decisions. This paper reports work in progress on learning decision policies in the…

机器学习 · 计算机科学 2020-11-04 Dennis Wei
‹ 上一页 1 8 9 10 下一页 ›