English
Related papers

Related papers: On the Convergence of Min-Max Langevin Dynamics an…

200 papers

Markov chains and diffusion processes are indispensable tools in machine learning and statistics that are used for inference, sampling, and modeling. With the growth of large-scale datasets, the computational cost associated with simulating…

Statistics Theory · Mathematics 2017-08-31 Jonathan H. Huggins , James Zou

In order to solve tasks like uncertainty quantification or hypothesis tests in Bayesian imaging inverse problems, we often have to draw samples from the arising posterior distribution. For the usually log-concave but high-dimensional…

Computation · Statistics 2025-01-23 Matthias J. Ehrhardt , Lorenz Kuger , Carola-Bibiane Schönlieb

We present the first general bounds on the mixing time of the Markov chain associated to the logit dynamics for wide classes of strategic games. The logit dynamics with inverse noise beta describes the behavior of a complex system whose…

Computer Science and Game Theory · Computer Science 2012-12-12 Vincenzo Auletta , Diodato Ferraioli , Francesco Pasquale , Paolo Penna , Giuseppe Persiano

Stochastic games are an important class of problems that generalize Markov decision processes to game theoretic scenarios. We consider finite state two-player zero-sum stochastic games over an infinite time horizon with discounted rewards.…

Optimization and Control · Mathematics 2008-06-17 Parikshit Shah , Pablo A. Parrilo

We derive various exact results for Markovian systems that spontaneously relax to a non-equilibrium steady-state by using joint probability distributions symmetries of different entropy production decompositions. The analytical approach is…

Statistical Mechanics · Physics 2012-02-10 Reinaldo García-García , Vivien Lecomte , A. B. Kolton , D. Domínguez

We examine online safe multi-agent reinforcement learning using constrained Markov games in which agents compete by maximizing their expected total rewards under a constraint on expected total utilities. Our focus is confined to an episodic…

Machine Learning · Computer Science 2023-06-02 Dongsheng Ding , Xiaohan Wei , Zhuoran Yang , Zhaoran Wang , Mihailo R. Jovanović

A major goal in Algorithmic Game Theory is to justify equilibrium concepts from an algorithmic and complexity perspective. One appealing approach is to identify natural distributed algorithms that converge quickly to an equilibrium. This…

Computer Science and Game Theory · Computer Science 2018-06-14 Yun Kuen Cheung , Richard Cole , Yixin Tao

We review convergence and behavior of stochastic gradient descent for convex and nonconvex optimization, establishing various conditions for convergence to zero of the variance of the gradient of the objective function, and presenting a…

Optimization and Control · Mathematics 2025-03-06 Kevin Buck , Jessica Babyak , Paolo Piersanti , Kevin Zumbrun , Christiane Gallos , Dorothea Gallos

In this paper, we study a regularised relaxed optimal control problem and, in particular, we are concerned with the case where the control variable is of large dimension. We introduce a system of mean-field Langevin equations, the invariant…

Probability · Mathematics 2019-10-07 Kaitong Hu , Anna Kazeykina , Zhenjie Ren

We study the problem of finding the Nash equilibrium in a two-player zero-sum Markov game. Due to its formulation as a minimax optimization program, a natural approach to solve the problem is to perform gradient descent/ascent with respect…

Optimization and Control · Mathematics 2022-10-13 Sihan Zeng , Thinh T. Doan , Justin Romberg

We study how risk-sensitive players act in situations where the outcome is influenced not only by the state-action profile but also by the distribution of it. In such interactive decision-making problems, the classical mean-field game…

Optimization and Control · Mathematics 2015-05-26 Hamidou Tembine

We present a novel method for drawing samples from Gibbs distributions with densities of the form $\pi(x) \propto \exp(-U(x))$. The method accelerates the unadjusted Langevin algorithm by introducing an inertia term similar to Polyak's…

Numerical Analysis · Mathematics 2025-10-09 Alexander Falk , Andreas Habring , Christoph Griesbacher , Thomas Pock

We consider zero-sum stochastic games for continuous time Markov decision processes with risk-sensitive average cost criterion. Here the transition and cost rates may be unbounded. We prove the existence of the value of the game and a…

Optimization and Control · Mathematics 2021-09-21 Mrinal K. Ghosh , Subrata Golui , Chandan Pal , Somnath Pradhan

We introduce a continuous policy-value iteration algorithm where the approximations of the value function of a stochastic control problem and the optimal control are simultaneously updated through Langevin-type dynamics. This framework…

Optimization and Control · Mathematics 2025-06-11 Qi Feng , Gu Wang

In this paper, we consider a continuous-type Bayesian Nash equilibrium (BNE) seeking problem in subnetwork zero-sum games, which is a generalization of deterministic subnetwork zero-sum games and discrete-type Bayesian zero-sum games. In…

Optimization and Control · Mathematics 2023-09-15 Hanzheng Zhang , Guanpu Chen , Yiguang Hong

A Dynkin game is considered for stochastic differential equations with random coefficients. We first apply Qiu and Tang's maximum principle for backward stochastic partial differential equations to generalize Krylov estimate for the…

Optimization and Control · Mathematics 2011-09-27 Shanjian Tang , Zhou Yang

We study dynamic finite-player and mean-field stochastic games within the framework of Markov perfect equilibria (MPE). Our focus is on discrete time and space structures without monotonicity. Unlike their continuous-time analogues,…

Optimization and Control · Mathematics 2025-09-29 Felix Höfer , H. Mete Soner , Atilla Yılmaz

We study best-response type learning dynamics for zero-sum polymatrix games under two information settings. The two settings are distinguished by the type of information that each player has about the game and their opponents' strategy. The…

Optimization and Control · Mathematics 2025-08-13 Fathima Zarin Faizal , Asuman Ozdaglar , Martin J. Wainwright

The Distributional Alignment Game framework provides a powerful variational perspective on Answer-Level Fine-Tuning (ALFT). However, standard algorithms for these games rely on estimating logarithmic rewards from small batches, introducing…

Machine Learning · Computer Science 2026-05-05 Mehryar Mohri , Jon Schneider , Yutao Zhong

Langevin dynamics has become a popular tool to simulate the Boltzmann equilibrium distribution. When the repartition of the Langevin equation involves the exact realization of the Ornstein-Uhlenbeck noise, in addition to the conventional…

Chemical Physics · Physics 2017-11-15 Dezhang Li , Xu Han , Yichen Chai , Cong Wang , Zifei Chen , Zhijun Zhang , Jian Liu , Jiushu Shao