English
Related papers

Related papers: On the policy improvement algorithm in continuous …

200 papers

In this paper, we demonstrate the existence of team-optimal strategies for static teams under observation-sharing information structures. Assuming that agents can access shared observations, we begin by converting the team problem into an…

Systems and Control · Electrical Eng. & Systems 2023-09-15 Naci Saldi

The connections between optimal control and Bayesian inference have long been recognised, with the field of stochastic (optimal) control combining these frameworks for the solution of partially observable control problems. In particular,…

Optimization and Control · Mathematics 2022-03-10 Manuel Baltieri

Solutions to address the periodic review inventory control problem with nonstationary random demand, lost sales, and stochastic vendor lead times typically involve making strong assumptions on the dynamics for either approximation or…

Machine Learning · Statistics 2023-10-26 Dean Foster , Randy Jia , Dhruv Madeka

Stochastic optimization naturally arises in machine learning. Efficient algorithms with provable guarantees, however, are still largely missing, when the objective function is nonconvex and the data points are dependent. This paper studies…

Machine Learning · Computer Science 2018-10-02 Minshuo Chen , Lin Yang , Mengdi Wang , Tuo Zhao

Cellular automata are capable of developing complex behaviors based on simple local interactions between their elements. Some of these characteristics have been used to propose and improve meta-heuristics for global optimization; however,…

Many reinforcement learning algorithms are built on an assumption that an agent interacts with an environment over fixed-duration, discrete time steps. However, physical systems are continuous in time, requiring a choice of…

Machine Learning · Computer Science 2024-09-04 Kris De Asis , Richard S. Sutton

We design receding horizon control strategies for stochastic discrete-time linear systems with additive (possibly) unbounded disturbances, while obeying hard bounds on the control inputs. We pose the problem of selecting an appropriate…

Optimization and Control · Mathematics 2011-07-07 Debasish Chatterjee , Peter Hokayem , John Lygeros

We present a pair of adjoint optimal control problems characterizing a class of time-symmetric stochastic processes defined on random time intervals. The associated PDEs are of free-boundary type. The particularity of our approach is that…

Probability · Mathematics 2020-07-07 Ana Bela Cruzeiro , Carlos Oliveira , Jean-Claude Zambrini

In this paper, we study the stochastic optimal control problem for control system with time-varying delay. The corresponding stochastic differential equation is a kind of stochastic differential delay equation. We prove the existence and…

Optimization and Control · Mathematics 2024-01-17 Yuhang Li , Yuecai Han

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

Optimization and Control · Mathematics 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang

Metric Temporal Logic can express temporally evolving properties with time-critical constraints or time-triggered constraints for real-time systems. This paper extends the Metric Interval Temporal Logic with a distribution eventuality…

Formal Languages and Automata Theory · Computer Science 2021-05-12 Lening Li , Jie Fu

In this paper, we study a data-enabled predictive control (DeePC) algorithm applied to unknown stochastic linear time-invariant systems. The algorithm uses noise-corrupted input/output data to predict future trajectories and compute optimal…

Optimization and Control · Mathematics 2019-11-04 Jeremy Coulson , John Lygeros , Florian Dörfler

Stochastic dynamical systems are fundamental in state estimation, system identification and control. System models are often provided in continuous time, while a major part of the applied theory is developed for discrete-time systems.…

Dynamical Systems · Mathematics 2014-02-07 Niklas Wahlström , Patrix Axelsson , Fredrik Gustafsson

Discrete-time stochastic systems with continuous spaces are hard to verify and control, even with MDP abstractions due to the curse of dimensionality. We propose an abstraction-based framework with robust dynamic programming mappings that…

Systems and Control · Electrical Eng. & Systems 2026-05-13 Ruohan Wang , Siyuan Liu , Zhiyong Sun , Sofie Haesaert

We consider a data-driven formulation of the classical discrete-time stochastic control problem. Our approach exploits the natural structure of many such problems, in which significant portions of the system are uncontrolled. Employing the…

Optimization and Control · Mathematics 2025-08-25 Boris Baros , Samuel N. Cohen , Christoph Reisinger

In this paper we present an information theoretic approach to stochastic optimal control problems for systems with compound Poisson noise. We generalize previous work on information theoretic path integral control to discontinuous dynamics…

Optimization and Control · Mathematics 2019-07-03 Ziyi Wang , Grady Williams , Evangelos A. Theodorou

Policy iteration is one of the classical frameworks of reinforcement learning, which requires a known initial stabilizing control. However, finding the initial stabilizing control depends on the known system model. To relax this requirement…

Systems and Control · Electrical Eng. & Systems 2025-03-20 Dongdong Li , Jiuxiang Dong

Folklore says that Howard's Policy Improvement Algorithm converges extraordinarily fast, even for controlled diffusion settings. In a previous paper, we proved that approximations of the solution of a particular parabolic partial…

Optimization and Control · Mathematics 2017-09-20 Jun Maeda , Saul D. Jacka

We introduce the framework of performative control, where the policy chosen by the controller affects the underlying dynamics of the control system. This results in a sequence of policy-dependent system state data with policy-dependent…

Optimization and Control · Mathematics 2024-10-31 Songfu Cai , Fei Han , Xuanyu Cao

We present an anytime algorithm which computes policies for decision problems represented as multi-stage influence diagrams. Our algorithm constructs policies incrementally, starting from a policy which makes no use of the available…

Artificial Intelligence · Computer Science 2013-02-01 Michael C. Horsch , David L. Poole
‹ Prev 1 4 5 6 7 8 10 Next ›