English
Related papers

Related papers: Thompson Sampling Algorithm for Stochastic Games

200 papers

We consider distributed learning problem in games with an unknown cost-relevant parameter, and aim to find the Nash equilibrium while learning the true parameter. Inspired by the social learning literature, we propose a distributed…

Optimization and Control · Mathematics 2023-03-14 Shijie Huang , Jinlong Lei , Yiguang Hong

This paper presents a new framework for analyzing and designing no-regret algorithms for dynamic (possibly adversarial) systems. The proposed framework generalizes the popular online convex optimization framework and extends it to its…

Machine Learning · Computer Science 2016-08-30 Ian Gemp , Sridhar Mahadevan

We design a distributed algorithm to seek generalized Nash equilibria of a robust game with uncertain coupled constraints. Due to the uncertainty of parameters in set constraints, we aim to find a generalized Nash equilibrium in the worst…

Optimization and Control · Mathematics 2022-04-05 Gehui Xu , Guanpu Chen , Hongsheng Qi

Correlated equilibrium generalizes Nash equilibrium by allowing a central coordinator to guide players' actions through shared recommendations, similar to how routing apps guide drivers. We investigate how a coordinator can learn a…

Computer Science and Game Theory · Computer Science 2025-09-16 Zhenlong Fang , Aryan Deshwal , Yue Yu

We introduce an online learning algorithm in the bandit feedback model that, once adopted by all agents of a congestion game, results in game-dynamics that converge to an $\epsilon$-approximate Nash Equilibrium in a polynomial number of…

Computer Science and Game Theory · Computer Science 2024-01-19 Leello Dadi , Ioannis Panageas , Stratis Skoulakis , Luca Viano , Volkan Cevher

This paper develops a predictive compensation framework for finite-horizon, discrete-time linear quadratic dynamic games subject to Gauss-Markov execution deviations from feedback Nash strategies. One player's control is corrupted by…

Systems and Control · Electrical Eng. & Systems 2025-11-18 Navid Mojahed , Mahdis Rabbani , Shima Nazari

In this paper we study the nonzero-sum Dynkin game in continuous time which is a two player non-cooperative game on stopping times. We show that it has a Nash equilibrium point for general stochastic processes. As an application, we…

Pricing of Securities · Quantitative Finance 2008-12-10 Said Hamadene , Jianfeng Zhang

We investigate properties of Thompson Sampling in the stochastic multi-armed bandit problem with delayed feedback. In a setting with i.i.d delays, we establish to our knowledge the first regret bounds for Thompson Sampling with arbitrary…

Machine Learning · Computer Science 2022-05-24 Han Wu , Stefan Wager

In this paper, we consider the problem of finding a Nash equilibrium in a multi-player game over generally connected networks. This model differs from a conventional setting in that players have partial information on the actions of their…

Optimization and Control · Mathematics 2019-12-10 Farzad Salehisadaghiani , Wei Shi , Lacra Pavel

We investigate a multi-agent decision problem in population games where each agent in a population makes a decision on strategy selection and revision to engage in repeated games with others. The strategy revision is subject to time delays…

Systems and Control · Electrical Eng. & Systems 2023-04-12 Shinkyu Park

This paper studies the limits of empirical means of open-loop Nash equilibria of linear-quadratic stochastic differential games as the number of players goes to infinity, when the corresponding mean field game is of potential type and may…

Probability · Mathematics 2026-02-27 Alekos Cecchin , Jodi Dianetti

A fundamental open problem in monotone game theory is the computation of a specific generalized Nash equilibrium (GNE) among all the available ones, e.g. the optimal equilibrium with respect to a system-level objective. The existing GNE…

Systems and Control · Electrical Eng. & Systems 2022-03-16 Emilio Benenati , Wicak Ananduta , Sergio Grammatico

In this paper, we study sequential decision-making for maximizing the Sharpe ratio (SR) in a stochastic multi-armed bandit (MAB) setting. Unlike standard bandit formulations that maximize cumulative reward, SR optimization requires…

Machine Learning · Computer Science 2026-04-02 Mohammad Taha Shah , Sabrina Khurshid , Gourab Ghatak

Motivated by economic applications such as recommender systems, we study the behavior of stochastic bandits algorithms under \emph{strategic behavior} conducted by rational actors, i.e., the arms. Each arm is a \emph{self-interested}…

Machine Learning · Computer Science 2020-11-16 Zhe Feng , David C. Parkes , Haifeng Xu

We study the problem of no-regret learning algorithms for general monotone and smooth games and their last-iterate convergence properties. Specifically, we investigate the problem under bandit feedback and strongly uncoupled dynamics, which…

Computer Science and Game Theory · Computer Science 2024-08-19 Jing Dong , Baoxiang Wang , Yaoliang Yu

This article is related to risk-sensitive nonzero-sum stochastic differential games in the Markovian framework. This game takes into account the attitudes of the players toward risk and the utility is of exponential form. We show the…

Optimization and Control · Mathematics 2014-12-04 Said Hamadène , Rui Mu

This paper studies a class of strongly monotone games involving non-cooperative agents that optimize their own time-varying cost functions. We assume that the agents can observe other agents' historical actions and choose actions that best…

Optimization and Control · Mathematics 2023-09-04 Zifan Wang , Yi Shen , Michael M. Zavlanos , Karl H. Johansson

Stochastic differential games have been used extensively to model agents' competitions in Finance, for instance, in P2P lending platforms from the Fintech industry, the banking system for systemic risk, and insurance markets. The recently…

Optimization and Control · Mathematics 2021-03-23 Jiequn Han , Ruimeng Hu , Jihao Long

Bayesian optimization through Gaussian process regression is an effective method of optimizing an unknown function for which every measurement is expensive. It approximates the objective function and then recommends a new measurement point…

Machine Learning · Statistics 2017-05-17 Hildo Bijl , Thomas B. Schön , Jan-Willem van Wingerden , Michel Verhaegen

Through a stochastic control theoretic approach, we analyze reputation games where a strategic long-lived player acts in a sequential repeated game against a collection of short-lived players. The key assumption in our model is that the…

Optimization and Control · Mathematics 2020-01-22 Nuh Aygün Dalkıran , Serdar Yüksel
‹ Prev 1 8 9 10 Next ›