English
Related papers

Related papers: Yet again on iteration improvement for averaged ex…

200 papers

We study reinforcement learning for controlled diffusion processes with unbounded continuous state spaces, bounded continuous actions, and polynomially growing rewards: settings that arise naturally in finance, economics, and operations…

Machine Learning · Computer Science 2025-12-18 Hanqing Jin , Renyuan Xu , Yanzhao Yang

We consider the adaptive test for the parameter change in discretely observed ergodic diffusion processes based on the cusum test. Using two test statistics based on the two quasi-log likelihood functions of the diffusion parameter and the…

Statistics Theory · Mathematics 2020-04-30 Yozo Tonaki , Yusuke Kaino , Masayuki Uchida

We consider a classical stochastic control problem in which a diffusion process is controlled by a withdrawal process up to a termination time. The objective is to maximize the expected discounted value of the withdrawals until the…

Probability · Mathematics 2024-06-19 Hélène Guérin , Dante Mata , Jean-François Renaud , Alexandre Roch

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

Optimization and Control · Mathematics 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang

This work is the third part of a program initiated in arXiv:2111.13258, arXiv:2302.06571 aiming at the development of an intrinsic geometric well-posedness theory for Hamilton-Jacobi equations related to controlled gradient flow problems in…

Analysis of PDEs · Mathematics 2024-02-02 Giovanni Conforti , Richard C. Kraaij , Luca Tamanini , Daniela Tonon

This paper examines a Markovian model for the optimal irreversible investment problem of a firm aiming at minimizing total expected costs of production. We model market uncertainty and the cost of investment per unit of production capacity…

Probability · Mathematics 2017-01-10 Tiziano De Angelis , Salvatore Federico , Giorgio Ferrari

This paper is concerned with cost optimization of an insurance company. The surplus of the insurance company is modeled by a controlled regime switching diffusion, where the regime switching mechanism provides the fluctuations of the random…

Optimization and Control · Mathematics 2016-08-02 Chao Zhu

This note re-visits the rolling-horizon control approach to the problem of a Markov decision process (MDP) with infinite-horizon discounted expected reward criterion. Distinguished from the classical value-iteration approach, we develop an…

Optimization and Control · Mathematics 2022-06-07 Hyeong Soo Chang

Heterogeneous diffusion processes can be well described by an overdamped Langevin equation with space-dependent diffusivity $D(x)$. We investigate the ergodic and non-ergodic behavior of these processes in an arbitrary potential well $U(x)$…

Statistical Mechanics · Physics 2019-05-01 Xudong Wang , Weihua Deng , Yao Chen

Controllable diffusion generation often relies on various heuristics that are seemingly disconnected without a unified understanding. We bridge this gap with Diffusion Controller (DiffCon), a unified control-theoretic view that casts…

Machine Learning · Computer Science 2026-03-10 Tong Yang , Moonkyung Ryu , Chih-Wei Hsu , Guy Tennenholtz , Yuejie Chi , Craig Boutilier , Bo Dai

A general method is proposed which allows one to estimate drift and diffusion coefficients of a stochastic process governed by a Langevin equation. It extends a previously devised approach [R. Friedrich et al., Physics Letters A 271, 217…

Data Analysis, Statistics and Probability · Physics 2009-11-11 D. Kleinhans , R. Friedrich , A. Nawroth , J. Peinke

In this paper, we obtain the exact controllability for a refined stochastic wave equation with three controls by establishing a novel Carleman estimate for a backward hyperbolic-like operator. Compared with the known result, the novelty of…

Optimization and Control · Mathematics 2023-09-21 Zhonghua Liao , Qi Lü

This paper is devoted to an optimal control problem of fully coupled forward-backward stochastic differential equations driven by sub-diffusion, whose solutions are not Markov processes. The stochastic maximum principle is obtained, where…

Optimization and Control · Mathematics 2025-03-11 Chenhui Hao , Jingtao Shi , Shuaiqi Zhang

This article studies a portfolio optimization problem, where the market consisting of several stocks is modeled by a multi-dimensional jump-diffusion process with age-dependent semi-Markov modulated coefficients. We study risk sensitive…

Portfolio Management · Quantitative Finance 2019-10-21 Milan Kumar Das , Anindya Goswami , Nimit Rana

This article investigates the exact controllability of three-dimensional stochastic Maxwell equations, a coupled system comprising two stochastic partial differential equations. The research establishes the observability inequality for the…

Optimization and Control · Mathematics 2026-05-26 Liying Sun , Xiaohan Wang , Yongyi Yu

In this paper, we consider risk-sensitive discounted control problem for continuous-time jump Markov processes taking values in general state space. The transition rates of underlying continuous-time jump Markov processes and the cost rates…

Optimization and Control · Mathematics 2021-04-27 Chandan Pal , Subrata Golui

This paper presents experiments for embedded cooperative distributed model predictive control applied to a team of hovercraft floating on an air hockey table. The hovercraft collectively solve a centralized optimal control problem in each…

Robotics · Computer Science 2025-03-19 Gösta Stomberg , Roland Schwan , Andrea Grillo , Colin N. Jones , Timm Faulwasser

Stabilization of a class of time-varying parabolic equations with uncertain input data using Receding Horizon Control (RHC) is investigated. The diffusion coefficient and the initial function are prescribed as random fields. We consider…

Optimization and Control · Mathematics 2023-02-03 Behzad Azmi , Lukas Herrmann , Karl Kunisch

It has recently been shown that there are substantial differences in the regularity behavior of the empirical process based on scalar diffusions as compared to the classical empirical process, due to the existence of diffusion local time.…

Probability · Mathematics 2011-05-25 Angelika Rohde , Claudia Strauch

This paper establishes that an MDP with a unique optimal policy and ergodic associated transition matrix ensures the convergence of various versions of the Value Iteration algorithm at a geometric rate that exceeds the discount factor…

Machine Learning · Computer Science 2024-06-17 Arsenii Mustafin , Alex Olshevsky , Ioannis Ch. Paschalidis