中文
相关论文

相关论文: Relative Value Iteration for Stochastic Differenti…

200 篇论文

In this contribution, we derive ILEG, an iterative algorithm to find risk sensitive solutions to nonlinear, stochastic optimal control problems. The algorithm is based on a linear quadratic approximation of an exponential risk sensitive…

系统与控制 · 计算机科学 2015-12-23 Farbod Farshidian , Jonas Buchli

This paper presents a learning dynamic with almost sure convergence guarantee for any stochastic game with turn-based controllers (on state transitions) as long as stage-payoffs induce a zero-sum or identical-interest game. Stage-payoffs…

计算机科学与博弈论 · 计算机科学 2023-10-11 Muhammed O. Sayin

This paper studies a class of stationary mean-field games of singular stochastic control with regime-switching. The representative agent adjusts the dynamics of a Markov-modulated It\^o-diffusion via a two-sided singular stochastic control…

最优化与控制 · 数学 2024-12-31 Jodi Dianetti , Giorgio Ferrari , Ioannis Tzouanas

In this paper we study zero-sum two-player stochastic differential games with the help of theory of Backward Stochastic Differential Equations (BSDEs). At the one hand we generalize the results of the pioneer work of Fleming and Souganidis…

概率论 · 数学 2011-02-19 Rainer Buckdahn , Juan Li

Zero-sum stochastic games generalize the notion of Markov Decision Processes (i.e. controlled Markov chains, or stochastic dynamic programming) to the 2-player competitive case : two players jointly control the evolution of a state…

最优化与控制 · 数学 2019-05-17 Jérôme Renault

In decision-dependent games, multiple players optimize their decisions under a data distribution that shifts with their joint actions, creating complex dynamics in applications like market pricing. A practical consequence of these dynamics…

计算机科学与博弈论 · 计算机科学 2025-09-04 Guangzheng Zhong , Yang Liu , Jiming Liu

This paper is concerned with the stochastic linear quadratic Stackelberg differential game with overlapping information, where the diffusion terms contain the control and state variables. Here the term "overlapping" means that there are…

最优化与控制 · 数学 2018-05-01 Jingtao Shi , Guangchen Wang , Jie Xiong

We study a class of stochastic target games where one player tries to find a strategy such that the state process almost-surely reaches a given target, no matter which action is chosen by the opponent. Our main result is a geometric dynamic…

概率论 · 数学 2015-02-03 Bruno Bouchard , Marcel Nutz

The optimal value computation for turned-based stochastic games with reachability objectives, also known as simple stochastic games, is one of the few problems in $NP \cap coNP$ which are not known to be in $P$. However, there are some…

计算复杂性 · 计算机科学 2014-08-10 David Auger , Pierre COUCHENEY , Yann Strozecki

This paper deals with N-person nonzero-sum discrete-time Markov games under a probability criterion, in which the transition probabilities and reward functions are allowed to vary with time. Differing from the existing works on the expected…

概率论 · 数学 2025-05-16 Xin Guo , Xin Wen

We consider a convexity constrained Hamilton-Jacobi-Bellman-type obstacle problem for the value function of a zero-sum differential game with asymmetric information. We propose a convexity-preserving probabilistic numerical scheme for the…

数值分析 · 数学 2021-03-26 Ľubomír Baňas , Giorgio Ferrari , Tsiry A. Randrianasolo

We consider a general type of non-Markovian impulse control problems under adverse non-linear expectation or, more specifically, the zero-sum game problem where the adversary player decides the probability measure. We show that the upper…

最优化与控制 · 数学 2022-06-30 Magnus Perninge

Control problems not admitting the dynamic programming principle are known as time-inconsistent. The game-theoretic approach is to interpret such problems as intrapersonal dynamic games and look for subgame perfect Nash equilibria. A…

最优化与控制 · 数学 2020-05-04 Kristoffer Lindensjö

In this paper we study the zero-sum and nonzero-sum differential games with not assuming Isaacs condition. Along with the partition $\pi$ of the time interval $[0,T]$, we choose the suitable random non-anticipative strategy with delay to…

最优化与控制 · 数学 2015-07-20 Juan Li , Wenqiang Li

Static reduction of information structures (ISs) is a method that is commonly adopted in stochastic control, team theory, and game theory. One approach entails change of measure arguments, which has been crucial for stochastic analysis and…

最优化与控制 · 数学 2023-07-13 Sina Sanjari , Tamer Başar , Serdar Yüksel

In this paper, the known deterministic linear-quadratic Stackelberg game is revisited, whose open-loop Stackelberg solution actually possesses the nature of time inconsistency. To handle this time inconsistency, {a two-tier game framework…

最优化与控制 · 数学 2022-03-09 Yuan-Hua Ni , Liping Liu , Xinzhen Zhang

The properties of value functions of time inhomogeneous optimal stopping problem and zero-sum game (Dynkin game) are studied through time dependent Dirichlet form. Under the absolute continuity condition on the transition function of the…

最优化与控制 · 数学 2013-06-28 Yipeng Yang

This article is dedicated to the study of mixed zero-sum two-player stochastic differential games in the situation when the player's cost functionals are modeled by doubly controlled reflected backward stochastic equations with two barriers…

最优化与控制 · 数学 2013-07-30 Said Hamadene , Eduard Rotenstein , Adrian Zalinescu

We characterize the optimal control for a class of singular stochastic control problems as the unique solution to a related Skorokhod reflection problem. The considered optimization problems concern the minimization of a discounted cost…

最优化与控制 · 数学 2023-05-22 Jodi Dianetti , Giorgio Ferrari

We propose a new algorithm for a broad class of periodic time-varying Stochastic Game-Theoretic Riccati Differential Equations arising in Zero-Sum Linear-Quadratic Stochastic Differential Games. The algorithm is constructed via dual-layer…

数值分析 · 数学 2025-11-06 Yiyuan Wang
‹ 上一页 1 8 9 10 下一页 ›