English
Related papers

Related papers: Coupling and a generalised Policy Iteration Algori…

200 papers

This paper is concerned with a time-inconsistent stochastic optimal control problem in an infinite time horizon with a non-degenerate diffusion in the state equation. A major assumption is that people become rational after a large time.…

Optimization and Control · Mathematics 2025-09-19 Qingmeng Wei , Jiongmin Yong

Diffusion models have achieved huge empirical success in data generation tasks. Recently, some efforts have been made to adapt the framework of diffusion models to discrete state space, providing a more natural approach for modeling…

Machine Learning · Statistics 2024-02-15 Hongrui Chen , Lexing Ying

In this paper, we study the optimal dividend problem under the continuous time diffusion model with the bounded dividend rate from the Reinforcement Learning (RL) perspective. Unlike the standard literature, our main focus will be on…

Optimization and Control · Mathematics 2026-03-30 Lihua Bai , Thejani Gamage , Jin Ma , Gaozhan Wang

The problem of optimal stopping with finite horizon in discrete time is considered in view of maximizing the expected gain. The algorithm proposed in this paper is completely nonparametric in the sense that it uses observed data from the…

Statistics Theory · Mathematics 2013-07-24 Michael Kohler , Harro Walk

To solve distributed optimization efficiently with various constraints and nonsmooth functions, we propose a distributed mirror descent algorithm with embedded Bregman damping, as a generalization of conventional distributed…

Optimization and Control · Mathematics 2021-08-30 Guanpu Chen , Weijian Li , Gehui Xu , Yiguang Hong

This paper establishes that an MDP with a unique optimal policy and ergodic associated transition matrix ensures the convergence of various versions of the Value Iteration algorithm at a geometric rate that exceeds the discount factor…

Machine Learning · Computer Science 2024-06-17 Arsenii Mustafin , Alex Olshevsky , Ioannis Ch. Paschalidis

We study optimal stopping for diffusion processes with unknown model primitives within the continuous-time reinforcement learning (RL) framework developed by Wang et al. (2020), and present applications to option pricing and portfolio…

Optimization and Control · Mathematics 2025-08-12 Min Dai , Yu Sun , Zuo Quan Xu , Xun Yu Zhou

This paper analyzes single-item continuous-review inventory models with random supplies in which the inventory dynamic between orders is described by a diffusion process, and a long-term average cost criterion is used to evaluate decisions.…

Optimization and Control · Mathematics 2024-02-07 K. L. Helmes , R. H. Stockbridge , C. Zhu

We investigate an optimal investment problem with a general performance criterion which, in particular, includes discontinuous functions. Prices are modeled as diffusions and the market is incomplete. We find an explicit solution for the…

Probability · Mathematics 2008-12-02 Nikolai Dokuchaev , Ulrich Haussmann

We present a dynamic programming-based solution to a stochastic optimal control problem up to a hitting time for a discrete-time Markov control process. Firstly, we determine an optimal control policy to steer the process toward a compact…

Optimization and Control · Mathematics 2009-09-28 Debasish Chatterjee , Eugenio Cinquemani , Giorgos Chaloulos , John Lygeros

In this paper we discuss policy iteration methods for approximate solution of a finite-state discounted Markov decision problem, with a focus on feature-based aggregation methods and their connection with deep reinforcement learning…

Machine Learning · Computer Science 2018-08-23 Dimitri P. Bertsekas

A novel distributed algorithm is proposed for finite-time converging to a feasible consensus solution satisfying global optimality to a certain accuracy of the distributed robust convex optimization problem (DRCO) subject to bounded…

Optimization and Control · Mathematics 2023-09-06 Xunhao Wu , Jun Fu

This article is concerned with stability and performance of controlled stochastic processes under receding horizon policies. We carry out a systematic study of methods to guarantee stability under receding horizon policies via appropriate…

Systems and Control · Computer Science 2017-11-27 Debasish Chatterjee , John Lygeros

We consider zero-sum stochastic games with finite state and action spaces, perfect information, mean payoff criteria, without any irreducibility assumption on the Markov chains associated to strategies (multichain games). The value of such…

Optimization and Control · Mathematics 2012-08-03 Marianne Akian , Jean Cochet-Terrasson , Sylvie Detournay , Stéphane Gaubert

We introduce a family of hybrid discretisations for the numerical approximation of optimal control problems governed by the equations of immiscible displacement in porous media. The proposed schemes are based on mixed and discontinuous…

Numerical Analysis · Mathematics 2019-05-02 S. Kumar , R. Ruiz Baier , R. Sandilya

Lloyd's algorithm is an iterative method that solves the quantization problem, i.e. the approximation of a target probability measure by a discrete one, and is particularly used in digital applications. This algorithm can be interpreted as…

Optimization and Control · Mathematics 2026-05-14 Léo Portales , Elsa Cazelles , Edouard Pauwels

Towards bridging classical optimal control and online learning, regret minimization has recently been proposed as a control design criterion. This competitive paradigm penalizes the loss relative to the optimal control actions chosen by a…

Systems and Control · Electrical Eng. & Systems 2023-06-27 Andrea Martin , Luca Furieri , Florian Dörfler , John Lygeros , Giancarlo Ferrari-Trecate

We consider the task of filtering a dynamic parameter evolving as a diffusion process, given data collected at discrete times from a likelihood which is conjugate to the marginal law of the diffusion, when a generic dual process on a…

Probability · Mathematics 2023-11-29 Guillaume Kon Kam King , Andrea Pandolfi , Marco Piretto , Matteo Ruggiero

The paper is a full version of the short presentation in \cite{amv17}. Ergodic control for one-dimensional controlled diffusion is tackled; both drift and diffusion coefficients may depend on a strategy which is assumed markovian. Ergodic…

Probability · Mathematics 2020-09-01 Svetlana Anulova , Hilmar Mai , Alexander Veretennikov

This paper introduces a general framework for iterative optimization algorithms and establishes under general assumptions that their convergence is asymptotically geometric. We also prove that under appropriate assumptions, the rate of…

Machine Learning · Statistics 2023-02-27 Randal Douc , Sylvain Le Corff