English
Related papers

Related papers: A Payoff-Based Policy Gradient Method in Stochasti…

200 papers

This work investigates continuous time stochastic differential games with a large number of players, whose costs and dynamics interact through the empirical distribution of both their states and their controls. The control processes are…

Probability · Mathematics 2022-02-22 Peng Luo , Ludovic Tangpi

Game theory is a very profound study on distributed decision-making behavior and has been extensively developed by many scholars. However, many existing works rely on certain strict assumptions such as knowing the opponent's private…

Computer Science and Game Theory · Computer Science 2020-04-21 Kuo Chun Tsai , Zhu Han

We solve the stochastic generalized Nash equilibrium (SGNE) problem in merely monotone games with expected value cost functions. Specifically, we present the first distributed SGNE seeking algorithm for monotone games that requires one…

Optimization and Control · Mathematics 2021-07-15 Barbara Franci , Sergio Grammatico

Graphon games have been introduced to study games with many players who interact through a weighted graph of interaction. By passing to the limit, a game with a continuum of players is obtained, in which the interactions are through a…

Optimization and Control · Mathematics 2024-04-02 Mathieu Laurière , Ludovic Tangpi , Xuchen Zhou

Learning processes in games explain how players grapple with one another in seeking an equilibrium. We study a natural model of learning based on individual gradients in two-player continuous games. In such games, the arguably natural…

Computer Science and Game Theory · Computer Science 2020-11-10 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff

Evolutionary game theory is a powerful mathematical framework to study how intelligent individuals adjust their strategies in collective interactions. It has been widely believed that it is impossible to unilaterally control players'…

Optimization and Control · Mathematics 2021-08-31 Renfei Tan , Qi Su , Bin Wu , Long Wang

In this work, we study stochastic non-cooperative games, where only noisy black-box function evaluations are available to estimate the cost function for each player. Since each player's cost function depends on both its own decision…

Computer Science and Game Theory · Computer Science 2025-11-18 Haidong Li , Anzhi Sheng , Yijie Peng , Long Wang

Policy gradient methods enjoy strong practical performance in numerous tasks in reinforcement learning. Their theoretical understanding in multiagent settings, however, remains limited, especially beyond two-player competitive and potential…

Computer Science and Game Theory · Computer Science 2023-12-22 Ioannis Anagnostides , Ioannis Panageas , Gabriele Farina , Tuomas Sandholm

We examine perfect information stochastic mean-payoff games - a class of games containing as special sub-classes the usual mean-payoff games and parity games. We show that deterministic memoryless strategies that are optimal for discounted…

Computer Science and Game Theory · Computer Science 2010-06-09 Hugo Gimbert , Wiesław Zielonka

We use techniques from the statistical mechanics of disordered systems to analyse the properties of Nash equilibria of bimatrix games with large random payoff matrices. By means of an annealed bound, we calculate their number and analyse…

Disordered Systems and Neural Networks · Physics 2009-10-31 Johannes Berg , Martin Weigt

This paper presents a payoff perturbation technique, introducing a strong convexity to players' payoff functions in games. This technique is specifically designed for first-order methods to achieve last-iterate convergence in games where…

Computer Science and Game Theory · Computer Science 2025-03-04 Kenshi Abe , Mitsuki Sakamoto , Kaito Ariu , Atsushi Iwasaki

Modifying the reward-biased maximum likelihood method originally proposed in the adaptive control literature, we propose novel learning algorithms to handle the explore-exploit trade-off in linear bandits problems as well as generalized…

Machine Learning · Computer Science 2020-10-09 Yu-Heng Hung , Ping-Chun Hsieh , Xi Liu , P. R. Kumar

An extensive literature in economics and social science addresses contests, in which players compete to outperform each other on some measurable criterion, often referred to as a player's score, or output. Players incur costs that are an…

Computer Science and Game Theory · Computer Science 2013-08-01 Leslie Ann Goldberg , Paul W. Goldberg , Piotr Krysta , Carmine Ventre

An important challenge in non-cooperative game theory is coordinating on a single (approximate) equilibrium from many possibilities - a challenge that becomes even more complex when players hold private information. Recommender mechanisms…

Computer Science and Game Theory · Computer Science 2025-05-30 Bengisu Guresti , Chongjie Zhang , Yevgeniy Vorobeychik

We study distributed algorithms for seeking a Nash equilibrium in a class of non-cooperative convex games with strongly monotone mappings. Each player has access to her own smooth local cost function and can communicate to her neighbors in…

Optimization and Control · Mathematics 2018-10-24 Tatiana Tatarenko , Wei Shi , Angelia Nedich

Towards characterizing the optimization landscape of games, this paper analyzes the stability of gradient-based dynamics near fixed points of two-player continuous games. We introduce the quadratic numerical range as a method to…

Computer Science and Game Theory · Computer Science 2021-01-15 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff

Gradient-based approaches to direct policy search in reinforcement learning have received much recent attention as a means to solve problems of partial observability and to avoid some of the problems associated with policy degradation in…

Artificial Intelligence · Computer Science 2019-11-18 Jonathan Baxter , Peter L. Bartlett

The goal in this paper is to approximate the Price of Stability (PoS) in stochastic Nash games using stochastic approximation (SA) schemes. PoS is amongst the most popular metrics in game theory and provides an avenue for estimating the…

Optimization and Control · Mathematics 2023-10-31 Afrooz Jalilzadeh , Farzad Yousefian , Mohammadjavad Ebrahimi

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterative policy improvement. In previous work, heuristic rules are…

Machine Learning · Computer Science 2020-09-15 Yuanyi Zhong , Yuan Zhou , Jian Peng

The strategy improvement algorithm for mean payoff games and parity games is a local improvement algorithm, just like the simplex algorithm for linear programs. Their similarity has turned out very useful: many lower bounds on running time…

Computer Science and Game Theory · Computer Science 2025-09-22 Matthew Maat
‹ Prev 1 4 5 6 7 8 10 Next ›