English
Related papers

Related papers: Proper Policies in Infinite-State Stochastic Short…

200 papers

A temporally abstract action, or an option, is specified by a policy and a termination condition: the policy guides option behavior, and the termination condition roughly determines its length. Generally, learning with longer options (like…

Artificial Intelligence · Computer Science 2017-12-05 Anna Harutyunyan , Peter Vrancx , Pierre-Luc Bacon , Doina Precup , Ann Nowe

This paper describes the structure of optimal policies for discounted periodic-review single-commodity total-cost inventory control problems with fixed ordering costs for finite and infinite horizons. There are known conditions in the…

Optimization and Control · Mathematics 2017-05-30 Eugene A. Feinberg , Yan Liang

This work proposes an optimal safe controller minimizing an infinite horizon cost functional subject to control barrier functions (CBFs) safety conditions. The constrained optimal control problem is reformulated as a minimization problem of…

Systems and Control · Electrical Eng. & Systems 2022-02-03 Hassan Almubarak , Evangelos A. Theodorou , Nader Sadegh

We consider a class of closed loop stochastic optimal control problems in finite time horizon, in which the cost is an expectation conditional on the event that the process has not exited a given bounded domain. An important difficulty is…

Optimization and Control · Mathematics 2019-12-19 Yves Achdou , Mathieu Laurière , Pierre-Louis Lions

We consider a model of optimal investment and consumption with both habit formation and partial observations in incomplete It\^{o} processes market. The investor chooses his consumption under the addictive habits constraint while only…

Portfolio Management · Quantitative Finance 2014-08-12 Xiang Yu

Markov decision problems are most commonly solved via dynamic programming. Another approach is Bellman residual minimization, which directly minimizes the squared Bellman residual objective function. However, compared to dynamic…

Machine Learning · Computer Science 2026-04-28 Donghwan Lee , Hyukjun Yang

This paper is concerned with a stochastic recursive optimal control problem with time delay, where the controlled system is described by a stochastic differential delayed equation (SDDE) and the cost functional is formulated as the solution…

Optimization and Control · Mathematics 2014-08-26 Jingtao Shi , Huanshui Zhang

This paper investigates a Hamilton-Jacobi (HJ) analysis to solve finite-horizon optimal control problems for high-dimensional systems. Although grid-based methods, such as the level-set method [1], numerically solve a general class of HJ…

Systems and Control · Electrical Eng. & Systems 2021-06-28 Donggun Lee , Claire J. Tomlin

In this paper, we study a stochastic recursive optimal control problem in which the system is governed by a functional forward-backward stochastic differential equation. Under standard assumptions, we establish the dynamic programming…

Probability · Mathematics 2013-01-03 Shaolin Ji , Shuzhen Yang

Using game and probability theories, I study the French popular game 421, a perfect information stochastic stage game. The problem is to find strategies maximizing the probability of some expected utility. I only solve a player's round…

Optimization and Control · Mathematics 2007-05-23 Pierre Albarede

In this paper we consider an energy storage optimization problem in finite time in a model with partial information that allows for a changing economic environment. The state process consists of the storage level controlled by the storage…

Mathematical Finance · Quantitative Finance 2016-06-21 Anton A. Shardin , Michaela Szölgyenyi

We consider a class of exit time stochastic control problems for diffusion processes with discounted criterion, where the controller can utilize a given amount of resource, called "fuel". In contrast to the vast majority of existing…

Optimization and Control · Mathematics 2015-01-30 Dmitry B. Rokhlin , Georgii Mironenko

This paper presents an inverse optimality method to solve the Hamilton-Jacobi-Bellman equation for a class of nonlinear problems for which the cost is quadratic and the dynamics are affine in the input. The method is inverse optimal because…

Optimization and Control · Mathematics 2011-10-11 Luis Rodrigues , Didier Henrion , Mehdi Abedinpour Fallah

From the Hamilton-Jacobi-Bellman equation for the value function we derive a non-linear partial differential equation for the optimal portfolio strategy (the dynamic control). The equation is general in the sense that it does not depend on…

Portfolio Management · Quantitative Finance 2013-11-20 Mads Nielsen

We propose an algorithm to calculate the exact solution for utility optimization problems on finite state spaces under a class of non-differentiable preferences. We prove that optimal strategies must lie on a discrete grid in the plane, and…

Pricing of Securities · Quantitative Finance 2018-10-01 Marcellino Gaudenzi , Michel Vellekoop

For a general entropy-regularized stochastic control problem on an infinite horizon, we prove that a policy iteration algorithm (PIA) converges to an optimal relaxed control. Contrary to the standard stochastic control literature, classical…

Optimization and Control · Mathematics 2026-05-14 Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

One often encounters the curse of dimensionality in the application of dynamic programming to determine optimal policies for controlled Markov chains. In this paper, we provide a method to construct sub-optimal policies along with a bound…

Systems and Control · Computer Science 2011-08-17 Myoungkuk Park , Krishnamoorthy Kalyanam , Swaroop Darbha , Phil Chandler , Meir Pachter

An optimal control problem is considered for a stochastic differential equation with the cost functional determined by a backward stochastic Volterra integral equation (BSVIE, for short). This kind of cost functional can cover the general…

Optimization and Control · Mathematics 2019-11-13 Hanxiao Wang , Jiongmin Yong

We consider infinite-horizon stationary $\gamma$-discounted Markov Decision Processes, for which it is known that there exists a stationary optimal policy. Using Value and Policy Iteration with some error $\epsilon$ at each iteration, it is…

Machine Learning · Computer Science 2012-11-30 Bruno Scherrer , Boris Lesner

We establish a linear programming formulation for the solution of joint chance constrained optimal control problems over finite time horizons. The joint chance constraint may represent an invariance, reachability or reach-avoid…

Optimization and Control · Mathematics 2024-05-21 Niklas Schmid , Marta Fochesato , Tobias Sutter , John Lygeros
‹ Prev 1 4 5 6 7 8 10 Next ›