中文
相关论文

相关论文: Relative Value Iteration for Stochastic Differenti…

200 篇论文

In this paper, we study an infinite horizon non-autonomous stochastic recursive differential game. To this end, we first establish well-posedness and stability results for BSDEs with a time-dependent discount factor and a possibly unbounded…

最优化与控制 · 数学 2026-05-14 Sheng Huang , Qingmeng Wei

In two-player zero-sum stochastic games, where two competing players make decisions under uncertainty, a pair of optimal strategies is traditionally described by Nash equilibrium and computed under the assumption that the players have…

最优化与控制 · 数学 2019-07-30 Yagiz Savas , Mohamadreza Ahmadi , Takashi Tanaka , Ufuk Topcu

We tackle a fundamental problem in empirical game-theoretic analysis (EGTA), that of learning equilibria of simulation-based games. Such games cannot be described in analytical form; instead, a black-box simulator can be queried to obtain…

计算机科学与博弈论 · 计算机科学 2019-06-03 Enrique Areyan Viqueira , Cyrus Cousins , Eli Upfal , Amy Greenwald

We introduce a new non-zero-sum game of optimal stopping with asymmetric exercise opportunities. Given a stochastic process modelling the value of an asset, one player observes and can act on the process continuously, while the other player…

概率论 · 数学 2024-05-16 José Luis Pérez , Neofytos Rodosthenous , Kazutoshi Yamazaki

We prove regularity and stochastic homogenization results for certain degenerate elliptic equations in nondivergence form. The equation is required to be strictly elliptic, but the ellipticity may oscillate on the microscopic scale and is…

偏微分方程分析 · 数学 2014-10-29 Scott N. Armstrong , Charles K. Smart

In ergodic singular stochastic control problems, a decision-maker can instantaneously adjust the evolution of a state variable using a control of bounded variation, with the goal of minimizing a long-term average cost functional. The cost…

最优化与控制 · 数学 2025-10-14 Alessandro Calvia , Federico Cannerozzi , Giorgio Ferrari

We investigate an infinite dimensional partial differential equation of Isaacs' type, which arises from a zero-sum differential game between two masses. The evolution of the two masses is described by a controlled transport/continuity…

最优化与控制 · 数学 2025-05-07 Fabio Bagagiolo , Rossana Capuani , Luciano Marzufero

We study a problem of optimal irreversible investment and emission reduction formulated as a nonzero-sum dynamic game between an investor with environmental preferences and a firm. The game is set in continuous time on an infinite-time…

数理金融 · 定量金融 2026-03-31 Tiziano De Angelis , Caio César Graciani Rodrigues , Peter Tankov

We consider a zero-sum stochastic differential controller-and-stopper game in which the state process is a controlled diffusion evolving in a multi-dimensional Euclidean space. In this game, the controller affects both the drift and the…

最优化与控制 · 数学 2013-01-15 Erhan Bayraktar , Yu-Jui Huang

We formulate a new class of two-person zero-sum differential games, in a stochastic setting, where a specification on a target terminal state distribution is imposed on the players. We address such added specification by introducing…

系统与控制 · 电气工程与系统科学 2019-09-13 Yongxin Chen , Tryphon T. Georgiou , Michele Pavon

In this note we extend to the random, stationary ergodic setting previous results of periodic homogenization for a particular family of nonlinear nonlocal "elliptic" equations with oscillatory coefficients. Such equations include, but are…

偏微分方程分析 · 数学 2012-09-11 Russell W. Schwab

We propose an implementable, neural network-based structure preserving probabilistic numerical approximation for a generalized obstacle problem describing the value of a zero-sum differential game of optimal stopping with asymmetric…

数值分析 · 数学 2025-01-28 Ľubomír Baňas , Giorgio Ferrari , Tsiry Avisoa Randrianasolo

We present a numerical approach to finding optimal trajectories for players in a multi-body, asset-guarding game with nonlinear dynamics and non-convex constraints. Using the Iterative Best Response (IBR) scheme, we solve for each player's…

系统与控制 · 电气工程与系统科学 2020-11-04 Emmanuel Sin , Murat Arcak , Douglas Philbrick , Peter Seiler

The problem of order execution is cast as a relative entropy-regularized robust optimal control problem in this article. The order execution agent's goal is to maximize an objective functional associated with his profit-and-loss of trading…

最优化与控制 · 数学 2024-09-11 Meng Wang , Tai-Ho Wang

We are interested in the convergence of the value of n-stage games as n goes to infinity and the existence of the uniform value in stochastic games with a general set of states and finite sets of actions where the transition is commutative.…

最优化与控制 · 数学 2016-04-22 Xavier Venel

We study the asymptotic value of a frequency-dependent zero-sum game with separable payoff following a differential approach. The stage payoffs in such games depend on the current actions and on a linear function of the frequency of actions…

最优化与控制 · 数学 2019-01-23 Joseph Abdou , Nikolaos Pnevmatikos

Bayesian regression games are a special class of two-player general-sum Bayesian games in which the learner is partially informed about the adversary's objective through a Bayesian prior. This formulation captures the uncertainty in regard…

机器学习 · 计算机科学 2021-10-04 Wenshuo Guo , Michael I. Jordan , Tianyi Lin

Last-iterate convergence has received extensive study in two player zero-sum games starting from bilinear, convex-concave up to settings that satisfy the MVI condition. Typical methods that exhibit last-iterate convergence for the…

计算机科学与博弈论 · 计算机科学 2023-10-05 Yi Feng , Hu Fu , Qun Hu , Ping Li , Ioannis Panageas , Bo Peng , Xiao Wang

This paper is concerned with a non-zero sum differential game problem of an anticipated forward-backward stochastic differential delayed equation under partial information. We establish a necessary maximum principle and sufficient…

最优化与控制 · 数学 2017-02-17 Yi Zhuang

We generalize the results of Fleming and Souganidis (1989) on zero sum stochastic differential games to the case when the controls are unbounded. We do this by proving a dynamic programming principle using a covering argument instead of…

最优化与控制 · 数学 2012-01-17 Erhan Bayraktar , Song Yao