中文
相关论文

相关论文: Probabilistic Framework of Howard's Policy Iterati…

200 篇论文

This paper aims to extend the BML method proposed in Wang et al. [22] to make it applicable to more general coupled nonlinear FBSDEs. We interpret BML from the fixed-point iteration perspective and show that optimizing BML is equivalent to…

最优化与控制 · 数学 2023-11-28 Yutian Wang , Yuan-Hua Ni , Xun Li

A modified Deep BSDE (backward differential equation) learning method with measurability loss, called Deep BSDE-ML method, is introduced in this paper to solve a kind of linear decoupled forward-backward stochastic differential equations…

最优化与控制 · 数学 2022-01-06 Yutian Wang , Yuan-Hua Ni

We demonstrate that backward stochastic differential equations (BSDE) may be reformulated as ordinary functional differential equations on certain path spaces. In this framework, neither It\^{o}'s integrals nor martingale representation…

概率论 · 数学 2012-11-20 Gechun Liang , Terry Lyons , Zhongmin Qian

Applications in quantitative finance such as optimal trade execution, risk management of options, and optimal asset allocation involve the solution of high dimensional and nonlinear Partial Differential Equations (PDEs). The connection…

机器学习 · 统计学 2019-10-28 Batuhan Güler , Alexis Laignelet , Panos Parpas

We propose a new multistep deep learning-based algorithm for the resolution of moderate to high dimensional nonlinear backward stochastic differential equations (BSDEs) and their corresponding parabolic partial differential equations (PDE).…

数值分析 · 数学 2023-08-29 Daniel Bussell , Camilo Andrés García-Trillos

We propose a new algorithm for solving parabolic partial differential equations (PDEs) and backward stochastic differential equations (BSDEs) in high dimension, by making an analogy between the BSDE and reinforcement learning with the…

数值分析 · 数学 2020-07-14 Weinan E , Jiequn Han , Arnulf Jentzen

In this work, we concern with the high order numerical methods for coupled forward-backward stochastic differential equations (FBSDEs). Based on the FBSDEs theory, we derive two reference ordinary differential equations (ODEs) from the…

数值分析 · 数学 2014-03-27 Weidong Zhao , Yu Fu , Tao Zhou

Optimal control problems are inherently hard to solve as the optimization must be performed simultaneously with updating the underlying system. Starting from an initial guess, Howard's policy improvement algorithm separates the step of…

最优化与控制 · 数学 2020-05-25 B. Kerimkulov , D. Šiška , Ł. Szpruch

Deterministic Markov Decision Processes (DMDPs) are a mathematical framework for decision-making where the outcomes and future possible actions are deterministically determined by the current action taken. DMDPs can be viewed as a finite…

人工智能 · 计算机科学 2025-06-17 Ali Asadi , Krishnendu Chatterjee , Jakob de Raaij

The recently proposed numerical algorithm, deep BSDE method, has shown remarkable performance in solving high-dimensional forward-backward stochastic differential equations (FBSDEs) and parabolic partial differential equations (PDEs). This…

概率论 · 数学 2022-03-10 Jiequn Han , Jihao Long

The theory of Forward-Backward Stochastic Differential Equations (FBSDEs) paves a way to probabilistic numerical methods for nonlinear parabolic PDEs. The majority of the results on the numerical methods for FBSDEs relies on the global…

概率论 · 数学 2016-07-25 Arnaud Lionnet , Gonçalo dos Reis , Lukasz Szpruch

We introduce the deep multi-FBSDE method for robust approximation of coupled forward-backward stochastic differential equations (FBSDEs), focusing on cases where the deep BSDE method of Han, Jentzen, and E (2018) fails to converge. To…

数值分析 · 数学 2025-06-03 Kristoffer Andersson , Adam Andersson , Cornelis W. Oosterlee

It is well-known that decision-making problems from stochastic control can be formulated by means of a forward-backward stochastic differential equation (FBSDE). Recently, the authors of Ji et al. 2022 proposed an efficient deep learning…

最优化与控制 · 数学 2024-08-01 Zhipeng Huang , Balint Negyesi , Cornelis W. Oosterlee

Bounded policy iteration is an approach to solving infinite-horizon POMDPs that represents policies as stochastic finite-state controllers and iteratively improves a controller by adjusting the parameters of each node using linear…

人工智能 · 计算机科学 2012-06-18 Eric A. Hansen

Our goal is to compute a policy that guarantees improved return over a baseline policy even when the available MDP model is inaccurate. The inaccurate model may be constructed, for example, by system identification techniques when the true…

最优化与控制 · 数学 2015-06-17 Yinlam Chow , Marek Petrik , Mohammad Ghavamzadeh

We propose a new deep learning algorithm for solving high-dimensional parabolic integro-differential equations (PIDEs) and forward-backward stochastic differential equations with jumps (FBSDEJs). This novel algorithm can be viewed as an…

数值分析 · 数学 2025-10-28 Wansheng Wang , Jiangtao Pan , Jie Wang , Zaijun Ye

Backward stochastic differential equation (BSDE) provides probabilistic solutions for a class of parabolic partial differential equations (PDEs). DeepBSDE and FBSNN are two deep learning approaches for solving high-dimensional PDEs through…

数值分析 · 数学 2026-04-29 Zhao Zhang , Zhuopeng Hou

We provide performance guarantees for a variant of simulation-based policy iteration for controlling Markov decision processes that involves the use of stochastic approximation algorithms along with state-of-the-art techniques that are…

机器学习 · 计算机科学 2022-10-17 Anna Winnicki , R. Srikant

In this work, we propose a new deep learning-based scheme for solving high dimensional nonlinear backward stochastic differential equations (BSDEs). The idea is to reformulate the problem as a global optimization, where the local loss…

数值分析 · 数学 2024-04-18 Lorenc Kapllani , Long Teng

Model-based reinforcement learning approaches leverage a forward dynamics model to support planning and decision making, which, however, may fail catastrophically if the model is inaccurate. Although there are several existing methods…

机器学习 · 计算机科学 2020-09-30 Hang Lai , Jian Shen , Weinan Zhang , Yong Yu
‹ 上一页 1 2 3 10 下一页 ›