中文
相关论文

相关论文: Sybil-proof Mechanisms in Query Incentive Networks

200 篇论文

We introduce a simple network model that is inspired by social information networks such as twitter. Agents are nodes, connecting to another agent by building a directed edge has a cost, and reaching other agents via short directed paths…

计算机科学与博弈论 · 计算机科学 2017-05-10 L. Elisa Celis , Aida S. Mousavifar

This paper studies the optimal mechanism to motivate effort in a dynamic principal-agent model without transfers. An agent is engaged in a task with uncertain future rewards and can quit at any time. The principal knows the reward and…

理论经济学 · 经济学 2026-01-16 Chang Liu

As AI systems become increasingly autonomous, aligning their decision-making to human preferences is essential. In domains like autonomous driving or robotics, it is impossible to write down the reward function representing these…

Equipping agents with the capacity to justify made decisions using supporting evidence represents a cornerstone of accountable decision-making. Furthermore, ensuring that justifications are in line with human expectations and societal norms…

机器学习 · 计算机科学 2024-02-27 Aleksa Sukovic , Goran Radanovic

A principal who values an object allocates it to one or more agents. Agents learn private information (signals) from an information designer about the allocation payoff to the principal. Monetary transfer is not available but the principal…

理论经济学 · 经济学 2022-10-31 Yi-Chun Chen , Gaoji Hu , Xiangqian Yang

Demand Response (DR) is a program designed to match supply and demand by modifying consumption profile. Some of these programs are based on economic incentives, in which, a user is paid to reduce his energy requirements according to an…

计算机科学与博弈论 · 计算机科学 2018-05-31 José Vuelvas , Fredy Ruiz , Giambattista Gruosso

In reinforcement learning (RL), agents continually interact with the environment and use the feedback to refine their behavior. To guide policy optimization, reward models are introduced as proxies of the desired objectives, such that when…

机器学习 · 计算机科学 2025-06-19 Rui Yu , Shenghua Wan , Yucen Wang , Chen-Xiao Gao , Le Gan , Zongzhang Zhang , De-Chuan Zhan

With the rapid advancement of large language models and vision-language models, employing large models as Web Agents has become essential for automated web interaction. However, training Web Agents with reinforcement learning faces critical…

As machine learning models become more capable, they have exhibited increased potential in solving complex tasks. One of the most promising directions uses deep reinforcement learning to train autonomous agents in computer network defense…

机器学习 · 计算机科学 2023-10-23 Elizabeth Bates , Vasilios Mavroudis , Chris Hicks

We consider a finite-horizon discrete-time dynamic system that is jointly controlled by two strategic agents. There is a system designer that has its own reward function but does not have direct control over the agents' actions. We consider…

系统与控制 · 电气工程与系统科学 2026-05-12 Renyan Sun , Ashutosh Nayyar

We study allocation problems without monetary transfers where agents have correlated types, i.e., hold private information about one another. Such peer information is relevant in various settings, including science funding, allocation of…

理论经济学 · 经济学 2025-03-21 Axel Niemeyer , Justus Preusser

We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all…

计算机科学与博弈论 · 计算机科学 2021-05-07 Gianluca Brero , Alon Eden , Matthias Gerstgrasser , David C. Parkes , Duncan Rheingans-Yoo

We study the problem of mechanism design for allocating a set of indivisible items among agents with private preferences on items. We are interested in such a mechanism that is strategyproof (where agents' best strategy is to report their…

计算机科学与博弈论 · 计算机科学 2024-08-05 Ankang Sun , Bo Chen

This paper investigates the use of intrinsic reward to guide exploration in multi-agent reinforcement learning. We discuss the challenges in applying intrinsic reward to multiple collaborative agents and demonstrate how unreliable reward…

人工智能 · 计算机科学 2019-06-06 Wendelin Böhmer , Tabish Rashid , Shimon Whiteson

Recent advances in reinforcement learning (RL) have significantly enhanced the agentic capabilities of large language models (LLMs). In long-term and multi-turn agent tasks, existing approaches driven solely by outcome rewards often suffer…

机器学习 · 计算机科学 2026-03-19 Yuxiang Ji , Ziyu Ma , Yong Wang , Guanhua Chen , Xiangxiang Chu , Liaoni Wu

In the on-line Explore and Exploit literature, central to Machine Learning, a central planner is faced with a set of alternatives, each yielding some unknown reward. The planner's goal is to learn the optimal alternative as soon as…

计算机科学与博弈论 · 计算机科学 2015-07-31 Gal Bahar , Rann Smorodinsky , Moshe Tennenholtz

Selecting influentials in networks against strategic manipulations has attracted many researchers' attention and it also has many practical applications. Here, we aim to select one or two influentials in terms of progeny (the influential…

计算机科学与博弈论 · 计算机科学 2023-06-14 Yuxin Zhao , Yao Zhang , Dengji Zhao

We propose an incentive mechanism for the sponsored content provider market in which the communication of users can be represented by a graph and the private information of the users is assumed to have a continuous distribution function.…

计算机科学与博弈论 · 计算机科学 2023-03-27 Mina Montazeri , Pegah Rokhforoz , Hamed Kebriaei , Olga Fink

We present a physics-inspired method for inferring dynamic rankings in directed temporal networks - networks in which each directed and timestamped edge reflects the outcome and timing of a pairwise interaction. The inferred ranking of each…

We consider a finite-horizon discrete-time dynamic system jointly controlled by a designer and one or more agents, where the designer can influence the agents' actions through selective information disclosure. At each time step, the…

系统与控制 · 电气工程与系统科学 2025-08-04 Renyan Sun , Ashutosh Nayyar