English
Related papers

Related papers: Policy iteration for the deterministic control pro…

200 papers

The question of knowing whether the policy Iteration algorithm (PI) for solving Markov Decision Processes (MDPs) has exponential or (strongly) polynomial complexity has attracted much attention in the last 50 years. Recently, Fearnley…

Computer Science and Game Theory · Computer Science 2011-08-19 Romain Hollanders , Jean-Charles Delvenne , Raphaël Jungers

We consider the setting of iterative learning control, or model-based policy learning in the presence of uncertain, time-varying dynamics. In this setting, we propose a new performance metric, planning regret, which replaces the standard…

Machine Learning · Computer Science 2021-03-01 Naman Agarwal , Elad Hazan , Anirudha Majumdar , Karan Singh

This research considers the ranking and selection with input uncertainty. The objective is to maximize the posterior probability of correctly selecting the best alternative under a fixed simulation budget, where each alternative is measured…

Optimization and Control · Mathematics 2023-05-15 Hui Xiao , Zhihong Wei

This paper investigates a singular stochastic control problem for a multi-dimensional regime-switching diffusion process confined in an unbounded domain. The objective is to maximize the total expected discounted rewards from exerting the…

Optimization and Control · Mathematics 2016-08-02 Qingshuo Song , Chao Zhu

In this work, we consider the problem of steering the first two moments of the uncertain state of a discrete time nonlinear stochastic system to prescribed goal quantities at a given final time. In principle, the latter problem can be…

Optimization and Control · Mathematics 2020-10-01 Efstathios Bakolas , Alexandros Tsolovikos

In this paper, we analyze the convergence of several discretize-then-optimize algorithms, based on either a second-order or a fourth-order finite difference discretization, for solving elliptic PDE-constrained optimization or optimal…

Numerical Analysis · Mathematics 2018-08-14 Jun Liu , Zhu Wang

For a general entropy-regularized stochastic control problem on an infinite horizon, we prove that a policy iteration algorithm (PIA) converges to an optimal relaxed control. Contrary to the standard stochastic control literature, classical…

Optimization and Control · Mathematics 2026-05-14 Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

We study a linear-quadratic optimal control problem involving a parabolic equation with fractional diffusion and Caputo fractional time derivative of orders $s \in (0,1)$ and $\gamma \in (0,1]$, respectively. The spatial fractional…

Optimization and Control · Mathematics 2015-04-02 Harbir Antil , Enrique Otarola , Abner J. Salgado

We consider a general family of nonlocal in space and time diffusion equations with space-time dependent diffusivity and prove convergence of finite difference schemes in the context of viscosity solutions under very mild conditions. The…

Numerical Analysis · Mathematics 2023-11-27 Félix del Teso , Łukasz Płociniczak

Computing the rate-distortion function for continuous sources is commonly regarded as a standard continuous optimization problem. When numerically addressing this problem, a typical approach involves discretizing the source space and…

Information Theory · Computer Science 2024-05-02 Lingyi Chen , Shitong Wu , Wenyi Zhang , Huihui Wu , Hao Wu

In this paper, a space-time discontinuous Galerkin finite element method for distributed optimal control problems governed by unsteady diffusion-convection-reaction equations with control constraints is studied. Time discretization is…

Optimization and Control · Mathematics 2013-08-09 Tuğba Akman , Bülent Karasözen

In this paper, optimal control problems governed by diffusion equations with Dirichlet and Neumann boundary conditions are investigated in the framework of the gradient discretisation method. Gradient schemes are defined for the optimality…

Numerical Analysis · Mathematics 2018-10-09 Jerome Droniou , Neela Nataraj , Devika Shylaja

This paper studies the adaptive optimal control problem for a class of linear time-delay systems described by delay differential equations (DDEs). A crucial strategy is to take advantage of recent developments in reinforcement learning and…

Systems and Control · Electrical Eng. & Systems 2022-10-04 Leilei Cui , Bo Pang , Zhong-Ping Jiang

The purpose of this work is to study an optimal control problem for a semilinear elliptic partial differential equation with a linear combination of Dirac measures as a forcing term; the control variable corresponds to the amplitude of such…

Optimization and Control · Mathematics 2023-07-04 Enrique Otarola

This work proposes a decision-making framework for partially observable systems in continuous time with discrete state and action spaces. As optimal decision-making becomes intractable for large state spaces we employ approximation methods…

Machine Learning · Computer Science 2024-03-01 Yannick Eich , Bastian Alt , Heinz Koeppl

We present discrete-time approximation of optimal control policies for infinite horizon discounted/ergodic control problems for controlled diffusions in $\Rd$\,. In particular, our objective is to show near optimality of optimal policies…

Optimization and Control · Mathematics 2025-02-11 Somnath Pradhan , Serdar Yuksel

We study a class of infinite-horizon impulse control problems with execution delay in discrete time. Using probabilistic methods, particularly the notion of the Snell envelope of processes, we construct an optimal strategy among all…

Optimization and Control · Mathematics 2025-01-22 Said Hamadène , Boualem Djehiche

The paper is devoted to the study of a new class of optimal control problems for nonsmooth dynamical systems governed by nonconvex discontinuous differential inclusions of the sweeping type with involving variable time into optimization. We…

Optimization and Control · Mathematics 2025-03-05 Tan H. Cao , Boris S. Mordukhovich , Dao Nguyen , Trang Nguyen , Nguyen N. Thieu

In this work, we present numerical analysis for a distributed optimal control problem, with box constraint on the control, governed by a subdiffusion equation which involves a fractional derivative of order $\alpha\in(0,1)$ in time. The…

Numerical Analysis · Mathematics 2017-12-22 Bangti Jin , Buyang Li , Zhi Zhou

We develop a novel iterative algorithm for locally optimal experimental design under constraints, like budget or performance constraints. It is an adaptive discretization algorithm. In every iteration, a discretized version of the…

Optimization and Control · Mathematics 2026-04-21 Jochen Schmid , Philipp Seufert , Jan Schwientek , Tobias Seidel , Karl-Heinz Küfer
‹ Prev 1 3 4 5 6 7 10 Next ›