中文
相关论文

相关论文: Stochastic dynamic programming under recursive Eps…

200 篇论文

This paper studies a robust continuous-time Markowitz portfolio selection pro\-blem where the model uncertainty carries on the covariance matrix of multiple risky assets. This problem is formulated into a min-max mean-variance problem over…

投资组合管理 · 定量金融 2017-03-14 Amine Ismail , Huyên Pham

A mathematical framework for Continuous Time Finance based on operator algebraic methods offers a new direct and entirely constructive perspective on the field and leads to new numerical analysis techniques. This is partly a review paper as…

概率论 · 数学 2009-09-29 Claudio Albanese

Discrete time stochastic optimal control problems and Markov decision processes (MDPs) are fundamental models for sequential decision-making under uncertainty and as such provide the mathematical framework underlying reinforcement learning…

最优化与控制 · 数学 2025-07-01 Arnulf Jentzen , Konrad Kleinberg , Thomas Kruse

In this paper we study the behavior of finite dimensional fixed point iterations, induced by discretization of a continuous fixed point iteration defined within a Banach space setting. We show that the difference between the discrete…

数值分析 · 数学 2017-06-29 Mario Amrein

In this paper, we study the delayed stochastic recursive optimal control problem with a non-Lipschitz generator, in which both the dynamics of the control system and the recursive cost functional depend on the past path segment of the state…

最优化与控制 · 数学 2023-12-27 Jiaqiang Wen , Zhen Wu , Qi Zhang

We study value-iteration (VI) algorithms for solving general (a.k.a. multichain) Markov decision processes (MDPs) under the average-reward criterion, a fundamental but theoretically challenging setting. Beyond the difficulties inherent to…

最优化与控制 · 数学 2026-04-23 Matthew Zurek , Yudong Chen

In many branches of engineering, Banach contraction mapping theorem is employed to establish the convergence of certain deterministic algorithms. Randomized versions of these algorithms have been developed that have proved useful in…

概率论 · 数学 2023-09-25 Abhishek Gupta , Rahul Jain , Peter Glynn

Merton portfolio management problem is studied in this paper within a stochastic volatility, non constant time discount rate, and power utility framework. This problem is time inconsistent and the way out of this predicament is to consider…

投资组合管理 · 定量金融 2024-02-09 Oumar Mbodji , Traian A. Pirvu

We introduce a framework to approximate a Markov Decision Process that stands on two pillars: state aggregation -- as the algorithmic infrastructure; and central-limit-theorem-type approximations -- as the mathematical underpinning of…

最优化与控制 · 数学 2021-04-13 Amy B. Z. Zhang , Itai Gurvich

Although average gain optimality is a commonly adopted performance measure in Markov Decision Processes (MDPs), it is often too asymptotic. Further incorporating measures of immediate losses leads to the hierarchy of bias optimalities, all…

机器学习 · 计算机科学 2025-10-16 Victor Boone , Adrienne Tuynman

We investigate the adaptive robust control framework for portfolio optimization and loss-based hedging under drift and volatility uncertainty. Adaptive robust problems offer many advantages but require handling a double optimization problem…

最优化与控制 · 数学 2020-05-06 Tao Chen , Michael Ludkovski

We consider an incomplete market with a nontradable stochastic factor and a continuous time investment problem with an optimality criterion based on monotone mean-variance preferences. We formulate it as a stochastic differential game…

投资组合管理 · 定量金融 2023-04-25 Jakub Trybuła , Dariusz Zawisza

In this paper, the mean-variance portfolio selection problem with Poisson jumps are studied, where the recursive utility is given by the solution to a backward stochastic differential equation with Poisson jumps. Both the maximum principle…

最优化与控制 · 数学 2025-12-02 Qiyue Zhang , Jingtao Shi

Under the expected total reward criterion, the optimal value of a finite-horizon Markov decision process can be determined by solving the Bellman equations. The equations were extended by D. J. White to processes with vector rewards in…

最优化与控制 · 数学 2023-08-28 Anas Mifrani

Standard stochastic control methods assume that the probability distribution of uncertain variables is available. Unfortunately, in practice, obtaining accurate distribution information is a challenging task. To resolve this issue, we…

最优化与控制 · 数学 2021-10-13 Insoon Yang

This paper is devoted to a study of robust fundamental theorems of asset pricing in discrete time and finite horizon settings. Uncertainty is modelled by a (possibly uncountable) family of price processes on the same probability space. Our…

数理金融 · 定量金融 2024-04-04 Huy N. Chau

Recursive stochastic algorithms have gained significant attention in the recent past due to data driven applications. Examples include stochastic gradient descent for solving large-scale optimization problems and empirical dynamic…

机器学习 · 计算机科学 2020-07-27 Abhishek Gupta , Hao Chen , Jianzong Pi , Gaurav Tendolkar

The analysis of Temporal Difference (TD) learning in the average-reward setting faces notable theoretical difficulties because the Bellman operator is not contractive with respect to any norm. This complicates standard analyses of…

机器学习 · 计算机科学 2026-05-05 Haoxing Tian , Zaiwei Chen , Ioannis Ch. Paschalidis , Alex Olshevsky

In this paper we propose a new methodology for solving an uncertain stochastic Markovian control problem in discrete time. We call the proposed methodology the adaptive robust control. We demonstrate that the uncertain control problem under…

最优化与控制 · 数学 2017-06-08 Tomasz R. Bielecki , Tao Chen , Igor Cialenco , Areski Cousin , Monique Jeanblanc

We present the first finite-sample analysis of policy evaluation in robust average-reward Markov Decision Processes (MDPs). Prior work in this setting have established only asymptotic convergence guarantees, leaving open the question of…

机器学习 · 统计学 2025-12-11 Yang Xu , Washim Uddin Mondal , Vaneet Aggarwal