中文
相关论文

相关论文: Optimal Anonymous Independent Reward Scheme Design

200 篇论文

When users access shared resources in a selfish manner, the resulting societal cost and perceived users' cost is often higher than what would result from a centrally coordinated optimal allocation. While several contributions in mechanism…

计算机科学与博弈论 · 计算机科学 2024-03-08 Leonardo Pedroso , Andrea Agazzi , W. P. M. H. Heemels , Mauro Salazar

We investigate approximately optimal mechanisms in settings where bidders' utility functions are non-linear; specifically, convex, with respect to payments (such settings arise, for instance, in procurement auctions for energy). We provide…

计算机科学与博弈论 · 计算机科学 2017-02-23 Amy Greenwald , Takehiro Oyakawa , Vasilis Syrgkanis

In peer selection agents must choose a subset of themselves for an award or a prize. As agents are self-interested, we want to design algorithms that are impartial, so that an individual agent cannot affect their own chance of being…

计算机科学与博弈论 · 计算机科学 2020-05-01 Nicholas Mattei , Paolo Turrini , Stanislav Zhydkov

We present AIRS: Automatic Intrinsic Reward Shaping that intelligently and adaptively provides high-quality intrinsic rewards to enhance exploration in reinforcement learning (RL). More specifically, AIRS selects shaping function from a…

机器学习 · 计算机科学 2023-10-13 Mingqi Yuan , Bo Li , Xin Jin , Wenjun Zeng

We study the Regularized A-optimal Design (RAOD) problem, which selects a subset of $k$ experiments to minimize the inverse of the Fisher information matrix, regularized with a scaled identity matrix. RAOD has broad applications in Bayesian…

最优化与控制 · 数学 2025-05-22 Yongchun Li

The design of distributed algorithms is central to the study of multiagent systems control. In this paper, we consider a class of combinatorial cost-minimization problems and propose a framework for designing distributed algorithms with a…

系统与控制 · 计算机科学 2019-03-18 Rahul Chandan , Dario Paccagnan , Jason R. Marden

This paper presents a special type of distributed optimization problems, where the summation of agents' local cost functions (i.e., global cost function) is convex, but each individual can be non-convex. Unlike most distributed optimization…

最优化与控制 · 数学 2021-08-16 Yipeng Pang , Guoqiang Hu

In this paper, we investigate resource allocation algorithm design for intelligent reflecting surface (IRS)-assisted multiuser cognitive radio (CR) systems. In particular, an IRS is deployed to mitigate the interference caused by the…

信息论 · 计算机科学 2020-03-03 Dongfang Xu , Xianghao Yu , Robert Schober

We consider the problem of designing agents able to compute optimal decisions by composing data from multiple sources to tackle tasks involving: (i) tracking a desired behavior while minimizing an agent-specific cost; (ii) satisfying safety…

最优化与控制 · 数学 2023-05-23 Emiland Garrabe , Martina Lamberti , Giovanni Russo

When people choose routes minimizing their individual delay, the aggregate congestion can be much higher compared to that experienced by a centrally-imposed routing. Yet centralized routing is incompatible with the presence of…

系统与控制 · 电气工程与系统科学 2021-03-08 Mauro Salazar , Dario Paccagnan , Andrea Agazzi , W. P. M. H. , Heemels

Recent RL research has utilized reward shaping--particularly complex shaping rewards such as intrinsic motivation (IM)--to encourage agent exploration in sparse-reward environments. While often effective, ``reward hacking'' can lead to the…

机器学习 · 计算机科学 2025-05-20 Grant C. Forbes , Jianxun Wang , Leonardo Villalobos-Arias , Arnav Jhala , David L. Roberts

This paper designs a distributed stochastic annealing algorithm for non-convex cooperative aggregative games, whose agents' cost functions not only depend on agents' own decision variables but also rely on the sum of agents' decision…

最优化与控制 · 数学 2022-04-05 Yinghui Wang , Xiaoxue Geng , Guanpu Chen , Wenxiao Zhao

We consider a finite-horizon discrete-time dynamic system jointly controlled by a designer and one or more agents, where the designer can influence the agents' actions through selective information disclosure. At each time step, the…

系统与控制 · 电气工程与系统科学 2025-08-04 Renyan Sun , Ashutosh Nayyar

We consider the problem of assigning items to platforms where each item has a utility associated with each of the platforms to which it can be assigned. Each platform has a soft constraint over the total number of items it serves, modeled…

计算机科学与博弈论 · 计算机科学 2025-08-19 Atasi Panda , Harsh Sharma , Anand Louis , Prajakta Nimbhorkar

We apply control theoretic and optimization techniques to adaptively design incentives. In particular, we consider the problem of a planner with an objective that depends on data from strategic decision makers. The planner does not know the…

计算机科学与博弈论 · 计算机科学 2018-06-18 Lillian J. Ratliff , Tanner Fiez

Optimal design of experiments for correlated processes is an increasingly relevant and active research topic. Present methods have restricted possibilities to judge their quality. To fill this gap, we complement the virtual noise approach…

统计理论 · 数学 2021-10-25 Andrej Pázman , Markus Hainy , Werner G. Müller

Designing policies for a network of agents is typically done by formulating an optimization problem where each agent has access to state measurements of all the other agents in the network. Such policy designs with centralized information…

最优化与控制 · 数学 2024-05-02 Georgios Darivianakis , Angelos Georghiou , Soroosh Shafiee , John Lygeros

We introduce a dynamic mechanism design problem in which the designer wants to offer for sale an item to an agent, and another item to the same agent at some point in the future. The agent's joint distribution of valuations for the two…

计算机科学与博弈论 · 计算机科学 2023-05-22 Christos Papadimitriou , George Pierrakos , Christos-Alexandros Psomas , Aviad Rubinstein

Critical sectors of human society are progressing toward the adoption of powerful artificial intelligence (AI) agents, which are trained individually on behalf of self-interested principals but deployed in a shared environment. Short of…

多智能体系统 · 计算机科学 2021-12-22 Jiachen Yang , Ethan Wang , Rakshit Trivedi , Tuo Zhao , Hongyuan Zha

Autonomous agents must often deal with conflicting requirements, such as completing tasks using the least amount of time/energy, learning multiple tasks, or dealing with multiple opponents. In the context of reinforcement learning~(RL),…

机器学习 · 计算机科学 2019-10-30 Santiago Paternain , Luiz F. O. Chamon , Miguel Calvo-Fullana , Alejandro Ribeiro