中文
相关论文

相关论文: A Formal Solution to the Grain of Truth Problem

200 篇论文

We study the computational complexity of "public goods games on networks". In this model, each vertex in a graph is an agent that needs to take a binary decision of whether to "produce a good" or not. Each agent's utility depends on the…

计算机科学与博弈论 · 计算机科学 2022-07-12 Matan Gilboa , Noam Nisan

We study online Bayesian persuasion problems in which an informed sender repeatedly faces a receiver with the goal of influencing their behavior through the provision of payoff-relevant information. Previous works assume that the sender has…

计算机科学与博弈论 · 计算机科学 2024-11-12 Francesco Bacchiocchi , Matteo Bollini , Matteo Castiglioni , Alberto Marchesi , Nicola Gatti

We describe an algorithm for computing best response strategies in a class of two-player infinite games of incomplete information, defined by payoffs piecewise linear in agents' types and actions, conditional on linear comparisons of…

计算机科学与博弈论 · 计算机科学 2012-07-19 Daniel Reeves , Michael P. Wellman

Here, we develop a deep learning algorithm for solving Principal-Agent (PA) mean field games with market-clearing conditions -- a class of problems that have thus far not been studied and one that poses difficulties for standard numerical…

机器学习 · 计算机科学 2021-10-05 Steven Campbell , Yichao Chen , Arvind Shrivats , Sebastian Jaimungal

Many multi-agent interaction scenarios can be naturally modeled as noncooperative games, where each agent's decisions depend on others' future actions. However, deploying game-theoretic planners for autonomous decision-making requires a…

机器学习 · 计算机科学 2026-01-05 Yash Jain , Xinjie Liu , Lasse Peters , David Fridovich-Keil , Ufuk Topcu

It has been shown that one can accommodate data (Bayes) and constraints (MaxEnt) in one method, the method of Maximum (relative) Entropy (ME) (Giffin 2007). In this paper we show a complex agent based example of inference with two different…

统计方法学 · 统计学 2016-09-08 Adom Giffin

Distributed aggregative optimization methods are gaining increased traction due to their ability to address cooperative control and optimization problems, where the objective function of each agent depends not only on its own decision…

多智能体系统 · 计算机科学 2025-06-03 Ziqin Chen , Magnus Egerstedt , Yongqiang Wang

We introduce and study a computational version of the principal-agent problem -- a classic problem in Economics that arises when a principal desires to contract an agent to carry out some task, but has incomplete information about the agent…

计算机科学与博弈论 · 计算机科学 2023-05-18 David Hyland , Julian Gutierrez , Michael Wooldridge

We study the problem of selection in the context of Bayesian persuasion. We are given multiple agents with hidden values (or quality scores), to whom resources must be allocated by a welfare-maximizing decision-maker. An intermediary with…

计算机科学与博弈论 · 计算机科学 2025-11-18 Yannan Bai , Kamesh Munagala , Yiheng Shen , Davidson Zhu

We study the problem of eliciting and aggregating probabilistic information from multiple agents. In order to successfully aggregate the predictions of agents, the principal needs to elicit some notion of confidence from agents, capturing…

计算机科学与博弈论 · 计算机科学 2014-10-03 Rafael M. Frongillo , Yiling Chen , Ian A. Kash

The literature on strategic communication originated with the influential cheap talk model, which precedes the Bayesian persuasion model by three decades. This model describes an interaction between two agents: sender and receiver. The…

计算机科学与博弈论 · 计算机科学 2024-09-11 Yakov Babichenko , Inbal Talgam-Cohen , Haifeng Xu , Konstantin Zabarnyi

We consider a finite-horizon multi-armed bandit (MAB) problem in a Bayesian setting, for which we propose an information relaxation sampling framework. With this framework, we define an intuitive family of control policies that include…

机器学习 · 计算机科学 2021-06-17 Seungki Min , Costis Maglaras , Ciamac C. Moallemi

This paper addresses the problem of fair equilibrium selection in graphical games. Our approach is based on the data structure called the {\em best response policy}, which was proposed by Kearns et al. \cite{kls} as a way to represent all…

计算机科学与博弈论 · 计算机科学 2007-05-23 Edith Elkind , Leslie Ann Goldberg , Paul W. Goldberg

Existing multi-agent reinforcement learning methods are limited typically to a small number of agents. When the agent number increases largely, the learning becomes intractable due to the curse of the dimensionality and the exponential…

多智能体系统 · 计算机科学 2020-12-16 Yaodong Yang , Rui Luo , Minne Li , Ming Zhou , Weinan Zhang , Jun Wang

Methods for learning optimal policies in autonomous agents often assume that the way the domain is conceptualised---its possible states and actions and their causal structure---is known in advance and does not change during learning. This…

人工智能 · 计算机科学 2018-01-11 Craig Innes , Alex Lascarides , Stefano V Albrecht , Subramanian Ramamoorthy , Benjamin Rosman

Consider discrete-time linear distributed averaging dynamics, whereby agents in a network start with uncorrelated and unbiased noisy measurements of a common underlying parameter (state of the world) and iteratively update their estimates…

最优化与控制 · 数学 2023-03-21 Giacomo Como , Fabio Fagnani , Anton V. Proskurnikov

We study a dynamic model of Bayesian persuasion in sequential decision-making settings. An informed principal observes an external parameter of the world and advises an uninformed agent about actions to take over time. The agent takes…

计算机科学与博弈论 · 计算机科学 2022-05-25 Jiarui Gan , Rupak Majumdar , Goran Radanovic , Adish Singla

We study a setting in which a principal selects an agent to execute a collection of tasks according to a specified priority sequence. Agents, however, have their own individual priority sequences according to which they wish to execute the…

计算机科学与博弈论 · 计算机科学 2024-10-30 Donya G. Dobakhshari , Lav R. Varshney , Vijay Gupta

After experimenting with a number of non-probabilistic methods for dealing with uncertainty many researchers reaffirm a preference for probability methods [1] [2], although this remains controversial. The importance of being able to form…

人工智能 · 计算机科学 2013-04-11 Thomas Slack

If we could define the set of all bad outcomes, we could hard-code an agent which avoids them; however, in sufficiently complex environments, this is infeasible. We do not know of any general-purpose approaches in the literature to avoiding…

人工智能 · 计算机科学 2020-06-17 Michael K. Cohen , Marcus Hutter