中文
相关论文

相关论文: Empirical Centroid Fictitious Play: An Approach Fo…

200 篇论文

We propose a reinforcement learning algorithm for stationary mean-field games, where the goal is to learn a pair of mean-field state and stationary policy that constitutes the Nash equilibrium. When viewing the mean-field state and the…

机器学习 · 计算机科学 2020-10-12 Qiaomin Xie , Zhuoran Yang , Zhaoran Wang , Andreea Minca

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

This work proposes a novel distributed approach for computing a Nash equilibrium in convex games with merely monotone and restricted strongly monotone pseudo-gradients. By leveraging the idea of the centralized operator extrapolation method…

最优化与控制 · 数学 2025-07-18 Tatiana Tatarenko , Angelia Nedich

This work proposes a novel distributed approach for computing a Nash equilibrium in convex games with restricted strongly monotone pseudo-gradients. By leveraging the idea of the centralized operator extrapolation method presented in [4] to…

最优化与控制 · 数学 2023-10-25 Tatiana Tatarenko , Angelia Nedich

In this paper, a class of convex feasibility problems (CFPs) are studied for multi-agent systems through local interactions. The objective is to search a feasible solution to the convex inequalities with some set constraints in a…

系统与控制 · 计算机科学 2016-12-16 Kaihong Lu , Gangshan Jing , Long Wang

Federated learning aims to train predictive models for data that is distributed across clients, under the orchestration of a server. However, participating clients typically each hold data from a different distribution, which can yield to…

机器学习 · 计算机科学 2022-11-02 Sharut Gupta , Kartik Ahuja , Mohammad Havaei , Niladri Chatterjee , Yoshua Bengio

In this paper we introduce the novel framework of distributionally robust games. These are multi-player games where each player models the state of nature using a worst-case distribution, also called adversarial distribution. Thus each…

最优化与控制 · 数学 2017-07-25 Dario Bauso , Jian Gao , Hamidou Tembine

We present a simulation-based approach for solution of mean field games (MFGs), using the framework of empirical game-theoretical analysis (EGTA). Our primary method employs a version of the double oracle, iteratively adding strategies…

多智能体系统 · 计算机科学 2023-02-14 Yongzhao Wang , Michael P. Wellman

We consider learning by fictitious play in a large population of agents engaged in single-play, two-person rounds of a symmetric game, and derive a mean-filed type model for the corresponding stochastic process. Using this model, we…

计算机科学与博弈论 · 计算机科学 2019-01-11 Misha Perepelitsa

We propose a deep neural network-based algorithm to identify the Markovian Nash equilibrium of general large $N$-player stochastic differential games. Following the idea of fictitious play, we recast the $N$-player game into $N$ decoupled…

最优化与控制 · 数学 2020-06-08 Jiequn Han , Ruimeng Hu

High fidelity simulation of large-sized complex networks can be realized on a distributed computing platform that leverages the combined resources of multiple processors or machines. In a discrete event driven simulation, the assignment of…

分布式、并行与集群计算 · 计算机科学 2012-10-16 Aditya Kurve , Christopher Griffin , David J. Miller , George Kesidis

Constructing effective algorithms to converge to Nash Equilibrium (NE) is an important problem in algorithmic game theory. Prior research generally posits that the upper bound on the convergence rate for games is $O\left(T^{-1/2}\right)$.…

计算机科学与博弈论 · 计算机科学 2024-09-06 Qi Ju , Falin Hei , Yuxuan Liu , Zhemei Fang , Yunfeng Luo

This work considers stochastic differential games with a large number of players, whose costs and dynamics interact through the empirical distribution of both their states and their controls. We develop a new framework to prove convergence…

概率论 · 数学 2022-03-24 Mathieu Laurière , Ludovic Tangpi

Counterfactual Regret Minimization (CFR) and its variants developed based upon Regret Matching (RM) have been considered to be the best method to solve incomplete information extensive form games. In addition to RM and CFR, Fictitious Play…

计算机科学与博弈论 · 计算机科学 2023-11-14 Qi Ju

Federated learning aims to train predictive models for data that is distributed across clients, under the orchestration of a server. However, participating clients typically each hold data from a different distribution, whereby predictive…

机器学习 · 计算机科学 2022-05-24 Sharut Gupta , Kartik Ahuja , Mohammad Havaei , Niladri Chatterjee , Yoshua Bengio

We consider repeated games where the players behave according to cumulative prospect theory (CPT). We show that, when the players have calibrated strategies and behave according to CPT, the natural analog of the notion of correlated…

计算机科学与博弈论 · 计算机科学 2020-07-20 Soham R. Phade , Venkat Anantharam

We consider the problem of distributed channel allocation in large networks under the frequency-selective interference channel. Performance is measured by the weighted sum of achievable rates. Our proposed algorithm is a modified Fictitious…

信息论 · 计算机科学 2018-11-13 Ilai Bistritz , Amir Leshem

We investigate how well continuous-time fictitious play in two-player games performs in terms of average payoff, particularly compared to Nash equilibrium payoff. We show that in many games, fictitious play outperforms Nash equilibrium on…

计算机科学与博弈论 · 计算机科学 2014-11-20 Georg Ostrovski , Sebastian van Strien

We analyze the problem of distributed power allocation for orthogonal multiple access channels by considering a continuous non-cooperative game whose strategy space represents the users' distribution of transmission power over the network's…

计算机科学与博弈论 · 计算机科学 2015-03-19 Panayotis Mertikopoulos , Elena V. Belmega , Aris L. Moustakas , Samson Lasaulce

We focus on the problem of \emph{Answer-Level Fine-Tuning} (ALFT), where the goal is to optimize a language model based on the correctness or properties of its final answers, rather than the specific reasoning traces used to produce them.…

机器学习 · 计算机科学 2026-05-01 Mehryar Mohri , Jon Schneider , Yifan Wu