English
Related papers

Related papers: Dynamic Programming for Optimal Delivery Time Slot…

200 papers

We consider the value function of a stochastic optimal control of degenerate diffusion processes in a domain $D$. We study the smoothness of the value function, under the assumption of the non-degeneracy of the diffusion term along the…

Probability · Mathematics 2013-02-28 Wei Zhou

In this paper, we propose a general theory of ambiguity-averse MDPs, which treats the uncertain transition probabilities as random variables and evaluates a policy via a risk measure applied to its random return. This ambiguity-averse MDP…

Computer Science and Game Theory · Computer Science 2026-02-04 Axel Benyamine , Julien Grand-Clément , Marek Petrik , Michael I. Jordan , Alain Durmus

We construct an abstract framework in which the dynamic programming principle (DPP) can be readily proven. It encompasses a broad range of common stochastic control problems in the weak formulation, and deals with problems in the…

Optimization and Control · Mathematics 2019-06-04 Roman Fayvisovich , Gordan Zitkovic

In this paper, we study a stochastic optimal control problem under degenerate G-expectation. By using implied partition method, we show that the approximation result for admissible controls still hold. Based on this result, we prove that…

Optimization and Control · Mathematics 2022-10-19 Xiaojuan Li

In reinforcement learning, the value function is typically trained to solve the Bellman equation, which connects the current value to future values. This temporal dependency hints that the value function may contain implicit information…

Machine Learning · Computer Science 2025-01-17 Jacob Adamczyk

This paper presents a robust, distributed algorithm to solve general linear programs. The algorithm design builds on the characterization of the solutions of the linear program as saddle points of a modified Lagrangian function. We show…

Optimization and Control · Mathematics 2014-09-26 Dean Richert , Jorge Cortes

We study value-iteration (VI) algorithms for solving general (a.k.a. multichain) Markov decision processes (MDPs) under the average-reward criterion, a fundamental but theoretically challenging setting. Beyond the difficulties inherent to…

Optimization and Control · Mathematics 2026-04-23 Matthew Zurek , Yudong Chen

We consider a dynamic pricing problem where customer response to the current price is impacted by the customer price expectation, aka reference price. We study a simple and novel reference price mechanism where reference price is the…

Machine Learning · Computer Science 2024-07-23 Shipra Agrawal , Wei Tang

In this manuscript we consider a class optimal control problem for stochastic differential delay equations. First, we rewrite the problem in a suitable infinite-dimensional Hilbert space. Then, using the dynamic programming approach, we…

Optimization and Control · Mathematics 2023-02-20 Filippo de Feo , Salvatore Federico , Andrzej Święch

It is well known that mean-variance portfolio selection is a time-inconsistent optimal control problem in the sense that it does not satisfy Bellman's optimality principle and therefore the usual dynamic programming approach fails. We…

Portfolio Management · Quantitative Finance 2012-05-23 Christoph Czichowsky

The stochastic knapsack has been used as a model in wide ranging applications from dynamic resource allocation to admission control in telecommunication. In recent years, a variation of the model has become a basic tool in studying problems…

Pricing of Securities · Quantitative Finance 2008-12-02 Grace Lin , Yingdong Lu , David Yao

The problem of determining the European-style option price in the incomplete market has been examined within the framework of stochastic optimization. An analytic method based on the discrete dynamic programming equation (Bellman equation)…

Statistical Mechanics · Physics 2016-08-31 Sergei Fedotov , Sergei Mikhailov

In this paper, we consider the classic stochastic (dynamic) knapsack problem, a fundamental mathematical model in revenue management, with general time-varying random demand. Our main goal is to study the optimal policies, which can be…

Optimization and Control · Mathematics 2018-07-19 Yingdong Lu

This paper is devoted to the analysis of a finite horizon discrete-time stochastic optimal control problem, in presence of constraints. We study the regularity of the value function which comes from the dynamic programming algorithm. We…

Optimization and Control · Mathematics 2007-05-23 M. Papi , S. Sbaraglia

We study an open problem of risk-sensitive portfolio allocation in a regime-switching credit market with default contagion. The state space of the Markovian regime-switching process is assumed to be a countably infinite set. To characterize…

Portfolio Management · Quantitative Finance 2018-10-25 Lijun Bo , Huafu Liao , Xiang Yu

We provide a unifying approximate dynamic programming framework that applies to a broad variety of problems involving sequential estimation. We consider first the construction of surrogate cost functions for the purposes of optimization,…

Artificial Intelligence · Computer Science 2023-01-02 Dimitri Bertsekas

In this paper, we study an optimal stopping problem in the presence of model uncertainty and regime switching. The max-min formulation for robust control and the dynamic programming approach are adopted to establish a general theoretical…

Optimization and Control · Mathematics 2025-09-04 Siyu Lv , Zhen Wu , Jie Xiong , Xin Zhang

We study the dynamic pricing of discrete goods over a finite selling horizon. One way to capture both the elastic and stochastic reaction of purchases to price is through a model where sellers control the intensity of a counting process,…

Optimization and Control · Mathematics 2026-01-23 Burak Aydin , Emre Parmaksiz , Ronnie Sircar

In this paper we consider a family of optimal control problems for economic models whose state variables are driven by Delay Differential Equations (DDE's). We consider two main examples: an AK model with vintage capital and an advertising…

Optimization and Control · Mathematics 2007-05-23 Giorgio Fabbri , Silvia Faggian , Fausto Gozzi

We consider policy evaluation in infinite-horizon discounted Markov decision problems (MDPs) with infinite spaces. We reformulate this task a compositional stochastic program with a function-valued decision variable that belongs to a…

Optimization and Control · Mathematics 2020-05-19 Alec Koppel , Garrett Warnell , Ethan Stump , Peter Stone , Alejandro Ribeiro