中文
相关论文

相关论文: Asymmetric Information Acquisition Games

200 篇论文

This work presents a novel policy iteration algorithm to tackle nonzero-sum stochastic impulse games arising naturally in many applications. Despite the obvious impact of solving such problems, there are no suitable numerical methods…

最优化与控制 · 数学 2020-06-29 René Aïd , Francisco Bernal , Mohamed Mnif , Diego Zabaljauregui , Jorge P. Zubelli

This paper aims to design a distributed coordination algorithm for solving a multi-agent decision problem with a hierarchical structure. The primary goal is to search the Nash equilibrium of a noncooperative game such that each player has…

最优化与控制 · 数学 2022-05-17 Xiaoyu Ma , Jinlong Lei , Peng Yi , Jie Chen

A recent theory shows that a multi-player decentralized partially observable Markov decision process can be transformed into an equivalent single-player game, enabling the application of \citeauthor{bellman}'s principle of optimality to…

计算机科学与博弈论 · 计算机科学 2025-01-03 Johan Peralez , Aurélien Delage , Olivier Buffet , Jilles S. Dibangoye

We introduce a game model called "customer attraction game" to demonstrate the competition among online content providers. In this model, customers exhibit interest in various topics. Each content provider selects one topic and benefits…

计算机科学与博弈论 · 计算机科学 2024-10-11 Xiaotie Deng , Hangxin Gan , Ningyuan Li , Weian Li , Qi Qi

A recent body of experimental literature has studied empirical game-theoretical analysis, in which we have partial knowledge of a game, consisting of observations of a subset of the pure-strategy profiles and their associated payoffs to…

计算机科学与博弈论 · 计算机科学 2014-02-13 John Fearnley , Martin Gairing , Paul Goldberg , Rahul Savani

We develop a probabilistic approach to continuous-time finite state mean field games. Based on an alternative description of continuous-time Markov chain by means of semimartingale and the weak formulation of stochastic optimal control, our…

概率论 · 数学 2018-08-24 Rene Carmona , Peiqi Wang

In this paper, we propose an asynchronous distributed algorithm for the computation of generalized Nash equilibria in noncooperative games, where the players interact via an undirected communication graph. Specifically, we extend the paper…

计算机科学与博弈论 · 计算机科学 2024-12-20 Carlo Cenedese , Giuseppe Belgioioso , Sergio Grammatico , Ming Cao

There has been significant recent progress in algorithms for approximation of Nash equilibrium in large two-player zero-sum imperfect-information games and exact computation of Nash equilibrium in multiplayer strategic-form games. While…

计算机科学与博弈论 · 计算机科学 2025-10-01 Sam Ganzfried

This paper considers the maximization of information rates for the Gaussian frequency-selective interference channel, subject to power and spectral mask constraints on each link. To derive decentralized solutions that do not require any…

信息论 · 计算机科学 2008-01-17 Gesualdo Scutari , Daniel P. Palomar , Sergio Barbarossa

We study a very general class of games --- multi-dimensional aggregative games --- which in particular generalize both anonymous games and weighted congestion games. For any such game that is also large, we solve the equilibrium selection…

数据结构与算法 · 计算机科学 2015-02-26 Rachel Cummings , Michael Kearns , Aaron Roth , Zhiwei Steven Wu

We investigate the complexity of computing approximate Nash equilibria in anonymous games. Our main algorithmic result is the following: For any $n$-player anonymous game with a bounded number of strategies and any constant $\delta>0$, an…

计算机科学与博弈论 · 计算机科学 2016-08-29 Yu Cheng , Ilias Diakonikolas , Alistair Stewart

The distributed computation of Nash equilibria is assuming growing relevance in engineering where such problems emerge in the context of distributed control. Accordingly, we present schemes for computing equilibria of two classes of static…

最优化与控制 · 数学 2017-10-17 Hao Jiang , Uday V. Shanbhag , Sean P. Meyn

This paper studies the global Nash equilibrium problem of leader-follower multi-agent dynamics, which yields consensus with a privacy information encrypted learning algorithm. With the secure hierarchical structure, the relationship between…

系统与控制 · 电气工程与系统科学 2023-02-08 Kun Zhang , Ji-Feng Zhang , Rong Su , Huaguang Zhang

We study discrete-time mean-field Markov games with infinite numbers of agents where each agent aims to minimize its ergodic cost. We consider the setting where the agents have identical linear state transitions and quadratic cost…

最优化与控制 · 数学 2019-10-17 Zuyue Fu , Zhuoran Yang , Yongxin Chen , Zhaoran Wang

When learning in strategic environments, a key question is whether agents can overcome uncertainty about their preferences to achieve outcomes they could have achieved absent any uncertainty. Can they do this solely through interactions…

计算机科学与博弈论 · 计算机科学 2024-11-21 Nivasini Ananthakrishnan , Nika Haghtalab , Chara Podimata , Kunhe Yang

Claude Shannon's zero-error communication paradigm reshaped our understanding of fault-tolerant information transfer. Here, we adapt this notion into game theory with incomplete information. We ask: can players with private information…

In a situation where each player has control over the transition probabilities of each subsystem, we game-theoretically analyze the optimization problem of minimizing both the partial entropy production of each subsystem and a penalty for…

统计力学 · 物理学 2023-11-17 Yuma Fujimoto , Sosuke Ito

Although it has been known since the 1970s that a globally optimal strategy profile in a common-payoff game is a Nash equilibrium, global optimality is a strict requirement that limits the result's applicability. In this work, we show that…

计算机科学与博弈论 · 计算机科学 2022-07-08 Scott Emmons , Caspar Oesterheld , Andrew Critch , Vincent Conitzer , Stuart Russell

A wide variety of goals could cause an AI to disable its off switch because "you can't fetch the coffee if you're dead" (Russell 2019). Prior theoretical work on this shutdown problem assumes that humans know everything that AIs do. In…

计算机科学与博弈论 · 计算机科学 2024-12-10 Andrew Garber , Rohan Subramani , Linus Luu , Mark Bedaywi , Stuart Russell , Scott Emmons

We consider synthesis of control policies that maximize the probability of satisfying given temporal logic specifications in unknown, stochastic environments. We model the interaction between the system and its environment as a Markov…

系统与控制 · 计算机科学 2014-05-01 Jie Fu , Ufuk Topcu