English
Related papers

Related papers: On the policy improvement algorithm for ergodic ri…

200 papers

This paper considers time-inconsistent problems when control and stopping strategies are required to be made simultaneously (called stopping control problems by us). We first formulate the timeinconsistent stopping control problems under…

Optimization and Control · Mathematics 2023-06-21 Zongxia Liang , Fengyi Yuan

With the increasing pace of automation, modern robotic systems need to act in stochastic, non-stationary, partially observable environments. A range of algorithms for finding parameterized policies that optimize for long-term average…

Machine Learning · Computer Science 2019-09-04 David Nass , Boris Belousov , Jan Peters

This paper proposes a novel robust reinforcement learning framework for discrete-time linear systems with model mismatch that may arise from the sim-to-real gap. A key strategy is to invoke advanced techniques from control theory. Using the…

Systems and Control · Electrical Eng. & Systems 2023-12-07 Leilei Cui , Tamer Başar , Zhong-Ping Jiang

This paper aims to establish an entropy-regularized value-based reinforcement learning method that can ensure the monotonic improvement of policies at each policy update. Unlike previously proposed lower-bounds on policy improvement in…

Machine Learning · Computer Science 2020-08-26 Lingwei Zhu , Takamitsu Matsubara

The main challenge for adaptive regulation of linear-quadratic systems is the trade-off between identification and control. An adaptive policy needs to address both the estimation of unknown dynamics parameters (exploration), as well as the…

Systems and Control · Computer Science 2019-04-01 Mohamad Kazem Shirani Faradonbeh , Ambuj Tewari , George Michailidis

We consider minimizing the probability of falling below a target growth rate of the wealth process up to a time horizon $T$ in an incomplete market model, and then study the asymptotic behavior of minimizing probability as $T\to\infty$.…

Probability · Mathematics 2012-05-04 Hideo Nagai

We consider a classical stochastic control problem in which a diffusion process is controlled by a withdrawal process up to a termination time. The objective is to maximize the expected discounted value of the withdrawals until the…

Probability · Mathematics 2024-06-19 Hélène Guérin , Dante Mata , Jean-François Renaud , Alexandre Roch

This paper is concerned with a discounted stochastic optimal control problem for regime switching diffusion in an infinite horizon. First, as a preliminary with particular interests in its own right, the global well-posedness of infinite…

Optimization and Control · Mathematics 2026-02-06 Kai Ding , Xun Li , Siyu Lv , Xin Zhang

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

Optimization and Control · Mathematics 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang

We develop a neural-network framework for multi-period risk--reward stochastic control problems with constrained two-step feedback policies that may be discontinuous in the state. We allow a broad class of objectives built on a…

Computational Finance · Quantitative Finance 2026-03-09 Chang Chen , Duy-Minh Dang

In this work, we study the optimal discretization error of stochastic integrals, in the context of the hedging error in a multidimensional It\^{o} model when the discrete rebalancing dates are stopping times. We investigate the convergence,…

Probability · Mathematics 2014-05-19 Emmanuel Gobet , Nicolas Landon

This work is motivated by the need to study the impact of data uncertainties and material imperfections on the solution to optimal control problems constrained by partial differential equations. We consider a pathwise optimal control…

Optimization and Control · Mathematics 2016-03-01 Ahmad Ahmad Ali , Elisabeth Ullmann , Michael Hinze

Motivated by the applications, a class of optimal control problems is investigated, where the goal is to influence the behavior of a given population through another controlled one interacting with the first. Diffusive terms accounting for…

Optimization and Control · Mathematics 2023-03-10 Stefano Almi , Marco Morandotti , Francesco Solombrino

This paper analyzes the stability of optimal policies in the long-run stochastic control framework with an averaged risk-sensitive criterion for discrete-time MDPs on finite state-action space. In particular, we study the robustness of…

Optimization and Control · Mathematics 2025-09-23 Nicole Bäuerle , Marcin Pitera , Łukasz Stettner

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

General Economics · Economics 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang

Traditionally, systems governed by linear Partial Differential Equations (PDEs) are spatially discretized to exploit their algebraic structure and reduce the computational effort for controlling them. Due to beneficial insights of the PDEs,…

Systems and Control · Computer Science 2016-04-05 Saber Jafarizadeh

We analyze the stability of general nonlinear discrete-time stochastic systems controlled by optimal inputs that minimize an infinite-horizon discounted cost. Under a novel stochastic formulation of cost-controllability and detectability…

Optimization and Control · Mathematics 2025-04-30 Robert H. Moldenhauer , Dragan Nešić , Mathieu Granzotto , Romain Postoyan , Andrew R. Teel

This paper is to investigate the control problem of maximizing the net benefit of a single species while the cost of the resource allocation is minimized in a population model which can be described by a reaction diffusion advection…

Optimization and Control · Mathematics 2019-01-01 Lianzhang Bao , Huilai Li , Haojian Liang , Guangliang Zhao

We present a formulation of an optimal control problem for a two-dimensional diffusion process governed by a Fokker-Planck equation to achieve a nonequilibrium steady state with a desired circulation while accelerating convergence toward…

Systems and Control · Electrical Eng. & Systems 2026-03-26 Norihisa Namura , Hiroya Nakao

A novel distributed algorithm is proposed for finite-time converging to a feasible consensus solution satisfying global optimality to a certain accuracy of the distributed robust convex optimization problem (DRCO) subject to bounded…

Optimization and Control · Mathematics 2023-09-06 Xunhao Wu , Jun Fu
‹ Prev 1 4 5 6 7 8 10 Next ›