中文
相关论文

相关论文: Optimal Anonymous Independent Reward Scheme Design

200 篇论文

Synthesis of optimization algorithms typically follows a {\em design-then-analyze\/} approach, which can obscure fundamental performance limits and hinder the systematic development of algorithms that operate near these limits. Recently, a…

最优化与控制 · 数学 2025-09-26 Ibrahim K. Ozaslan , Wuwei Wu , Jie Chen , Tryphon T. Georgiou , Mihailo R. Jovanovic

Reinforcement learning (RL) is an effective approach to motion planning in autonomous driving, where an optimal driving policy can be automatically learned using the interaction data with the environment. Nevertheless, the reward function…

机器人学 · 计算机科学 2023-08-28 Lin-Chi Wu , Zengjie Zhang , Sofie Haesaert , Zhiqiang Ma , Zhiyong Sun

Teaming is the process of establishing connections among agents within a system to enable collaboration toward achieving a collective goal. This paper examines teaming in the context of a network of agents learning to coordinate with…

系统与控制 · 电气工程与系统科学 2025-04-08 Zhewei Wang , Olugbenga Moses Anubi , Marcos M. Vasconcelos

We propose a new approach to competitive analysis in online scheduling by introducing the novel concept of competitive-ratio approximation schemes. Such a scheme algorithmically constructs an online algorithm with a competitive ratio…

数据结构与算法 · 计算机科学 2012-11-01 Elisabeth Günther , Olaf Maurer , Nicole Megow , Andreas Wiese

A two-user downlink network aided by a reconfigurable intelligent surface is considered. The weighted sum signal to interference plus noise ratio maximization and the sum rate maximization models are presented, where the precoding vectors…

信号处理 · 电气工程与系统科学 2022-02-15 Cong Sun , Xian Liu , Bile Peng , Eduard Jorswieck

Entities in multi-agent systems may seek conflicting subobjectives, and this leads to competition between them. To address performance degradation due to competition, we consider a bi-level lottery where a social planner at the high level…

计算机科学与博弈论 · 计算机科学 2020-12-04 Hunmin Kim , Minghui Zhu

Many auction settings implicitly or explicitly require that bidders are treated equally ex-ante. This may be because discrimination is philosophically or legally impermissible, or because it is practically difficult to implement or…

计算机科学与博弈论 · 计算机科学 2014-11-06 Christos Tzamos , Christopher A. Wilkens

In this paper, we consider a general distributed system with multiple agents who select and then implement actions in the system. The system has an operator with a centralized objective. The agents, on the other hand, are selfinterested and…

计算机科学与博弈论 · 计算机科学 2020-01-15 Donya Ghavidel , Pratyush Chakraborty , Enrique Baeyens , Vijay Gupta , Pramod P. Khargonekar

In reinforcement learning (RL), different reward functions can define the same optimal policy but result in drastically different learning performance. For some, the agent gets stuck with a suboptimal behavior, and for others, it solves the…

机器学习 · 计算机科学 2025-02-25 Grigorii Veviurko , Wendelin Böhmer , Mathijs de Weerdt

There has been significant progress in deep reinforcement learning (RL) in recent years. Nevertheless, finding suitable hyperparameter configurations and reward functions remains challenging even for experts, and performance heavily relies…

机器学习 · 计算机科学 2024-10-10 Julian Dierkes , Emma Cramer , Holger H. Hoos , Sebastian Trimpe

Motivated by applications such as online labor markets we consider a variant of the stochastic multi-armed bandit problem where we have a collection of arms representing strategic agents with different performance characteristics. The…

计算机科学与博弈论 · 计算机科学 2025-03-11 Seyed A. Esmaeili , Suho Shin , Aleksandrs Slivkins

We show that computing the revenue-optimal deterministic auction in unit-demand single-buyer Bayesian settings, i.e. the optimal item-pricing, is computationally hard even in single-item settings where the buyer's value distribution is a…

计算机科学与博弈论 · 计算机科学 2015-03-10 Constantinos Daskalakis , Alan Deckelbaum , Christos Tzamos

Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal coordination, particularly when sparse task rewards provide insufficient grounding for coordinated…

机器学习 · 计算机科学 2026-04-07 Dogan Urgun , Gokhan Gungor

We consider the discrete assignment problem in which agents express ordinal preferences over objects and these objects are allocated to the agents in a fair manner. We use the stochastic dominance relation between fractional or randomized…

计算机科学与博弈论 · 计算机科学 2015-06-18 Haris Aziz , Serge Gaspers , Simon Mackenzie , Toby Walsh

In this paper, we consider a scenario of covert communication aided by multiple friendly interference nodes. The objective is to conceal the legitimate communication link under the surveillance of a warden. The main content is as follows:…

信号处理 · 电气工程与系统科学 2024-05-10 Xuyang Zhao. Wei Guo , Yongchao Wang

This paper investigates a two-stage game-theoretical model with multiple parallel rank-order contests. In this model, each contest designer sets up a contest and determines the prize structure within a fixed budget in the first stage.…

计算机科学与博弈论 · 计算机科学 2025-05-14 Xiaotie Deng , Ningyuan Li , Weian Li , Qi Qi

Collaborative machine learning (CML) provides a promising paradigm for democratizing advanced technologies by enabling cost-sharing among participants. However, the potential for rent-seeking behaviors among parties can undermine such…

机器学习 · 计算机科学 2025-01-03 Bingchen Wang , Zhaoxuan Wu , Fusheng Liu , Bryan Kian Hsiang Low

Machine learning and artificial intelligence conferences such as NeurIPS and ICML now regularly receive tens of thousands of submissions, posing significant challenges to maintaining the quality and consistency of the peer review process.…

机器学习 · 计算机科学 2026-01-23 Garrett G. Wen , Buxin Su , Natalie Collina , Zhun Deng , Weijie Su

I study the optimal design of ratings to motivate agent investment in quality when transfers are unavailable. The principal designs a rating scheme that maps the agent's quality to a (possibly stochastic) score. The agent has private…

理论经济学 · 经济学 2025-08-11 Peiran Xiao

This paper investigates an average secrecy rate (ASR) maximization problem for an unmanned aerial vehicle (UAV) enabled wireless communication system, wherein a UAV is employed to deliver confidential information to a ground destination in…

信息论 · 计算机科学 2021-02-23 Milad Tatar Mamaghani , Yi Hong
‹ 上一页 1 8 9 10 下一页 ›