English
Related papers

Related papers: Dynamic Programming for Optimal Delivery Time Slot…

200 papers

We consider a novel pricing and advertising framework, where a seller not only sets product price but also designs flexible 'advertising schemes' to influence customers' valuation of the product. We impose no structural restriction on the…

Computer Science and Game Theory · Computer Science 2024-12-12 Shipra Agrawal , Yiding Feng , Wei Tang

We present an arbitrage free theoretical framework for modeling bid and ask prices of dividend paying securities in a discrete time setup using theory of dynamic acceptability indices. In the first part of the paper we develop the theory of…

Pricing of Securities · Quantitative Finance 2014-12-31 Tomasz R. Bielecki , Igor Cialenco , Tao Chen

We consider a stochastic optimal control problem in a market model with temporary and permanent price impact, which is related to an expected utility maximization problem under finite fuel constraint. We establish the initial condition…

Mathematical Finance · Quantitative Finance 2015-10-13 Mourad Lazgham

This paper presents a new theory, known as robust dynamic pro- gramming, for a class of continuous-time dynamical systems. Different from traditional dynamic programming (DP) methods, this new theory serves as a fundamental tool to analyze…

Optimization and Control · Mathematics 2018-09-18 Tao Bian , Zhong-Ping Jiang

This paper aims to study the relationship between the maximum principle and the dynamic programming principle for recursive optimal control problem of stochastic evolution equations, where the control domain is not necessarily convex and…

Optimization and Control · Mathematics 2025-12-19 Ying Hu , Guomin Liu , Shanjian Tang

Linear Temporal Logic (LTL) is a formal way of specifying complex objectives for planning problems modeled as Markov Decision Processes (MDPs). The planning problem aims to find the optimal policy that maximizes the satisfaction probability…

Robotics · Computer Science 2024-08-13 Zetong Xuan , Yu Wang

In reinforcement learning (RL), aligning agent behavior with specific objectives typically requires careful design of the reward function, which can be challenging when the desired objectives are complex. In this work, we propose an…

Machine Learning · Computer Science 2025-09-05 Yuting Tang , Yivan Zhang , Johannes Ackermann , Yu-Jie Zhang , Soichiro Nishimori , Masashi Sugiyama

We study a dynamic stochastic control problem subject to Knightian uncertainty with multi-objective (vector-valued) criteria. Assuming the preferences across expected multi-loss vectors are represented by a given, yet general, preorder, we…

Optimization and Control · Mathematics 2024-07-02 Igor Cialenco , Gabriela Kováčová

In the Dynamic Programming approach to optimal control problems a crucial role is played by the value function that is characterized as the unique viscosity solution of a Hamilton-Jacobi-Bellman (HJB) equation. It is well known that this…

Numerical Analysis · Mathematics 2022-10-19 Luca Saluzzi , Alessandro Alla , Maurizio Falcone

In Bender and Dokuchaev (2013), we studied a control problem related to swing option pricing in a general non-Markovian setting. The main result there shows that the value process of this control problem can be uniquely characterized in…

Pricing of Securities · Quantitative Finance 2021-05-31 Christian Bender , Nikolai Dokuchaev

Value function estimation is an important task in reinforcement learning, i.e., prediction. The Boltzmann softmax operator is a natural value estimator and can provide several benefits. However, it does not satisfy the non-expansion…

Machine Learning · Computer Science 2019-09-10 Ling Pan , Qingpeng Cai , Qi Meng , Wei Chen , Longbo Huang , Tie-Yan Liu

In this paper, we study policy evaluation in continuous-time reinforcement learning (RL), where the state follows an unknown stochastic differential equation (SDE), but only discrete-time data are available. We first highlight that the…

Optimization and Control · Mathematics 2026-02-23 Yuhua Zhu

We consider the stochastic optimal control problem of McKean-Vlasov stochastic differential equation where the coefficients may depend upon the joint law of the state and control. By using feedback controls, we reformulate the problem into…

Probability · Mathematics 2017-03-09 Huyên Pham , Xiaoli Wei

In this work, we explore finite-dimensional linear representations of nonlinear dynamical systems by restricting the Koopman operator to an invariant subspace. The Koopman operator is an infinite-dimensional linear operator that evolves…

Dynamical Systems · Mathematics 2016-04-27 Steven L. Brunton , Bingni W. Brunton , Joshua L. Proctor , J. Nathan Kutz

Value functions derived from Markov decision processes arise as a central component of algorithms as well as performance metrics in many statistics and engineering applications of machine learning techniques. Computation of the solution to…

Machine Learning · Computer Science 2020-03-02 Adithya M. Devraj , Ioannis Kontoyiannis , Sean P. Meyn

We consider the problem of designing an expected-revenue maximizing mechanism for allocating multiple non-perishable goods of $k$ varieties to flexible consumers over $T$ time steps. In our model, a random number of goods of each variety…

Computer Science and Game Theory · Computer Science 2020-07-08 Shiva Navabi , Ashutosh Nayyar

This paper presents sufficient conditions for optimal control of systems with dynamics given by a linear operator, in order to obtain an explicit solution to the Bellman equation that can be calculated in a distributed fashion. Further, the…

Optimization and Control · Mathematics 2025-06-19 David Ohlin , Richard Pates , Murat Arcak

The paper is concerned with a variant of the continuous-time finite state Markov game of control and stopping where both players can affect transition rates, while only one player can choose a stopping time. We use the dynamic programming…

Optimization and Control · Mathematics 2022-08-09 Yurii Averboukh

We consider a general type of non-Markovian impulse control problems under adverse non-linear expectation or, more specifically, the zero-sum game problem where the adversary player decides the probability measure. We show that the upper…

Optimization and Control · Mathematics 2022-06-30 Magnus Perninge

In this paper, we aim to develop the theory of optimal stochastic control for branching diffusion processes where both the movement and the reproduction of the particles depend on the control. More precisely, we study the problem of…

Probability · Mathematics 2016-09-19 Julien Claisse