English
Related papers

Related papers: Inverse Dynamic Games Based on Maximum Entropy Inv…

200 papers

This letter studies multi-agent reinforcement learning in partially observable Markov potential games. Solving this problem is challenging due to partial observability, decentralized information, and the curse of dimensionality. First, to…

Multiagent Systems · Computer Science 2026-04-02 Wonseok Yang , Thinh T. Doan

Game theory studies situations in which strategic players can modify the state of a given system, due to the absence of a central authority. Solution concepts, such as Nash equilibrium, are defined to predict the outcome of such situations.…

Computer Science and Game Theory · Computer Science 2013-11-08 Diodato Ferraioli , Paul W. Goldberg , Carmine Ventre

Designing fair compensation mechanisms for demand response (DR) is challenging. This paper models the problem in a game theoretic setting and designs a payment distribution mechanism based on the Shapley Value. As exact computation of the…

Computer Science and Game Theory · Computer Science 2014-03-27 Gearóid O'Brien , Abbas El Gamal , Ram Rajagopal

The recent mean field game (MFG) formalism facilitates otherwise intractable computation of approximate Nash equilibria in many-agent settings. In this paper, we consider discrete-time finite MFGs subject to finite-horizon objectives. We…

Multiagent Systems · Computer Science 2022-07-11 Kai Cui , Heinz Koeppl

We study the inverse optimal control problem in social sciences: we aim at learning a user's true cost function from the observed temporal behavior. In contrast to traditional phenomenological works that aim to learn a generative model to…

Machine Learning · Computer Science 2018-05-23 Yichen Wang , Le Song , Hongyuan Zha

The gloabal objective of inverse Reinforcement Learning (IRL) is to estimate the unknown cost function of some MDP base on observed trajectories generated by (approximate) optimal policies. The classical approach consists in tuning this…

Machine Learning · Computer Science 2021-05-26 Firas Jarboui , Vianney Perchet

In many smart infrastructure applications flexibility in achieving sustainability goals can be gained by engaging end-users. However, these users often have heterogeneous preferences that are unknown to the decision-maker tasked with…

Computer Science and Game Theory · Computer Science 2017-04-27 Ioannis C. Konstantakopoulos , Lillian J. Ratliff , Ming Jin , S. Shankar Sastry , Costas Spanos

We consider a class of two-player dynamic stochastic nonzero-sum games where the state transition and observation equations are linear, and the primitive random variables are Gaussian. Each controller acquires possibly different dynamic…

Systems and Control · Computer Science 2014-01-21 Abhishek Gupta , Ashutosh Nayyar , Cedric Langbort , Tamer Basar

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterative policy improvement. In previous work, heuristic rules are…

Machine Learning · Computer Science 2020-09-15 Yuanyi Zhong , Yuan Zhou , Jian Peng

We consider a multi-player stochastic differential game with linear McKean-Vlasov dynamics and quadratic cost functional depending on the variance and mean of the state and control actions of the players in open-loop form. Finite and…

Probability · Mathematics 2018-12-04 Enzo Miller , Huyen Pham

Learning customer preferences from an observed behaviour is an important topic in the marketing literature. Structural models typically model forward-looking customers or firms as utility-maximizing agents whose utility is estimated using…

Computational Finance · Quantitative Finance 2017-12-14 Igor Halperin

This work studies the parameter identification problem of a generalized non-cooperative game, where each player's cost function is influenced by an observable signal and some unknown parameters. We consider the scenario where equilibrium of…

Computer Science and Game Theory · Computer Science 2023-10-17 Jianguo Chen , Jinlong Lei , Hongsheng Qi , Yiguang Hong

We study two-layer neural networks in the mean field limit, where the number of neurons tends to infinity. In this regime, the optimization over the neuron parameters becomes the optimization over the probability measures, and by adding an…

Optimization and Control · Mathematics 2023-08-17 Fan Chen , Zhenjie Ren , Songbo Wang

We investigate the convergence of symmetric stochastic differential games with interactions via control, where the volatility terms of both idiosyncratic and common noises are controlled. We apply the stochastic maximum principle, following…

Probability · Mathematics 2026-02-19 Erhan Bayraktar , Hiroaki Horikawa

In recent years, deep reinforcement learning has been shown to be adept at solving sequential decision processes with high-dimensional state spaces such as in the Atari games. Many reinforcement learning problems, however, involve…

Machine Learning · Computer Science 2018-06-05 Yiming Zhang , Quan Ho Vuong , Kenny Song , Xiao-Yue Gong , Keith W. Ross

This paper studies game-type credit default swaps that allow the protection buyer and seller to raise or reduce their respective positions once prior to default. This leads to the study of an optimal stopping game subject to early default…

Pricing of Securities · Quantitative Finance 2015-03-19 Masahiko Egami , Tim S. T. Leung , Kazutoshi Yamazaki

We use analytical techniques based on an expansion in the inverse system size to study the stochastic evolutionary dynamics of finite populations of players interacting in a repeated prisoner's dilemma game. We show that a mechanism of…

Populations and Evolution · Quantitative Biology 2012-04-20 Alex J. Bladon , Tobias Galla , Alan J. McKane

Despite its groundbreaking success, multi-agent reinforcement learning (MARL) still suffers from instability and nonstationarity. Replicator dynamics, the most well-known model from evolutionary game theory (EGT), provide a theoretical…

Machine Learning · Computer Science 2025-01-28 Tuo Zhang , Leonardo Stella , Julian Barreiro-Gomez

We study a continuous-time stochastic Stackelberg game in which a leader seeks to accomplish a primary objective while inferring a hidden parameter of a rational follower. The follower solves an entropy-regularized tracking problem and…

Optimization and Control · Mathematics 2025-10-08 Ruimeng Hu , Daniel Ralston , Xu Yang , Haosheng Zhou

Nash equilibria provide a principled framework for modeling interactions in multi-agent decision-making and control. However, many equilibrium-seeking methods implicitly assume that each agent has access to the other agents' objectives and…

Computer Science and Game Theory · Computer Science 2026-03-19 Mahdis Rabbani , Navid Mojahed , Shima Nazari