English
Related papers

Related papers: Policy Iteration Achieves Regularized Equilibrium …

200 papers

This article presents a constrained policy optimization approach for the optimal control of systems under nonstationary uncertainties. We introduce an assumption that we call Markov embeddability that allows us to cast the stochastic…

Optimization and Control · Mathematics 2026-05-11 Sungho Shin , François Pacaud , Emil Contantinescu , Mihai Anitescu

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

It is argued that a Gibbsian formula for the space-time distribution of microscopic trajectories of a nonequilibrium system provides a unifying framework for recent results on the fluctuations of the entropy production. The variable entropy…

Statistical Mechanics · Physics 2007-05-23 C. Maes

This paper is concerned with the open-loop time-consistent solution of time-inconsistent mean-field stochastic linear-quadratic optimal control. Different from standard stochastic linear-quadratic problems, both the system matrices and the…

Optimization and Control · Mathematics 2016-08-19 Yuan-Hua Ni , Ji-Feng Zhang , Miroslav Krstic

Generalising the idea of the classical EM algorithm that is widely used for computing maximum likelihood estimates, we propose an EM-Control (EM-C) algorithm for solving multi-period finite time horizon stochastic control problems. The new…

Economics · Quantitative Finance 2016-11-08 Steven Kou , Xianhua Peng , Xingbo Xu

Covariate balance is a conventional key diagnostic for methods used estimating causal effects from observational studies. Recently, there is an emerging interest in directly incorporating covariate balance in the estimation. We study a…

Methodology · Statistics 2017-02-14 Qingyuan Zhao , Daniel Percival

Equipping approximate dynamic programming (ADP) with inputconstraints has a tremendous significance. This enables ADP to be applied tothe systems with actuator limitations, which is quite common for dynamicalsystems. In a conventional…

Optimization and Control · Mathematics 2018-05-24 Xuefeng Bao , Zhi-Hong Mao , Nitin Sharma

This paper is concerned with the axiomatic foundation and explicit construction of a general class of optimality criteria that can be used for investment problems with multiple time horizons, or when the time horizon is not known in…

Portfolio Management · Quantitative Finance 2014-02-03 Sergey Nadtochiy , Michael Tehranchi

We present a policy iteration algorithm for the infinite-horizon N-player general-sum deterministic linear quadratic dynamic games and compare it to policy gradient methods. We demonstrate that the proposed policy iteration algorithm is…

Optimization and Control · Mathematics 2024-10-07 Yuxiang Guan , Giulio Salizzoni , Maryam Kamgarpour , Tyler H. Summers

Theoretical analyses of evolution strategies are indispensable for gaining a deep understanding of their inner workings. For constrained problems, rather simple problems are of interest in the current research. This work presents a…

Neural and Evolutionary Computing · Computer Science 2019-08-12 Patrick Spettel , Hans-Georg Beyer

The paper tests the validity of the critique of the fiscal theory of the price level. A stochastic general equilibrium model with continuous time is constructed. An active fiscal policy and a passive monetary policy have been set. Monetary…

Theoretical Economics · Economics 2024-03-05 Andrey Kofnov

This paper is devoted to solving a time-inconsistent risk-sensitive control problem with parameter $\e$ and its limit case ($\e\rightarrow0^+$) for countable-stated Markov decision processes (MDPs for short). Since the cost functional is…

Optimization and Control · Mathematics 2020-10-22 Hongwei Mei

We consider discrete-time infinite horizon deterministic optimal control problems with nonnegative cost per stage, and a destination that is cost-free and absorbing. The classical linear-quadratic regulator problem is a special case. Our…

Optimization and Control · Mathematics 2017-12-20 Dimitri P. Bertsekas

Recent studies have extended the use of the stochastic Hamilton-Jacobi-Bellman (HJB) equation to include complex variables for deriving quantum mechanical equations. However, these studies often assume that it is valid to apply the HJB…

Quantum Physics · Physics 2024-10-14 Vasil Yordanov

Nonzero-sum stochastic differential games with impulse controls offer a realistic and far-reaching modelling framework for applications within finance, energy markets, and other areas, but the difficulty in solving such problems has…

Numerical Analysis · Mathematics 2020-06-29 Diego Zabaljauregui

To integrate large systems of nonlinear differential equations in time, we consider a variant of nonlinear waveform relaxation (also known as dynamic iteration or Picard-Lindel\"of iteration), where at each iteration a linear inhomogeneous…

Numerical Analysis · Mathematics 2024-04-23 Mike A. Botchev

In this article, we study a continuous-time stochastic $H_\infty$ control problem based on reinforcement learning (RL) techniques that can be viewed as solving a stochastic linear-quadratic two-person zero-sum differential game (LQZSG).…

Optimization and Control · Mathematics 2024-10-02 Zhongshi Sun , Guangyan Jia

This work proposes a novel numerical scheme for solving the high-dimensional Hamilton-Jacobi-Bellman equation with a functional hierarchical tensor ansatz. We consider the setting of stochastic control, whereby one applies control to a…

Numerical Analysis · Mathematics 2025-07-01 Xun Tang , Nan Sheng , Lexing Ying

The electrical impedance tomography (EIT) problem of estimating the unknown conductivity distribution inside a domain from boundary current or voltage measurements requires the solution of a nonlinear inverse problem. Sparsity promoting…

Numerical Analysis · Mathematics 2024-05-27 Daniela Calvetti , Monica Pragliola , Erkki Somersalo

A time-inconsistent optimal control problem is formulated and studied for a controlled linear ordinary differential equation with quadratic cost functional. A notion of equilibrium control is introduced, which can be regarded as a…

Optimization and Control · Mathematics 2012-04-10 Jiongmin Yong
‹ Prev 1 8 9 10 Next ›