English
Related papers

Related papers: Mean field for Markov Decision Processes: from Dis…

200 papers

In this paper we propose a new way of proving the value of a firm that is currently producing a certain product and faces the option to exit the market. The problem of optimal exiting is an optimal stopping problem, that can be solved using…

Optimization and Control · Mathematics 2013-09-23 Manuel Guerra , Cláudia Nunes , Carlos Oliveira

In this paper, we present a mean field game to model the production behaviors of a very large number of producers, whose carbon emissions are regulated by government. Especially, an emission permits trading scheme is considered in our…

Economics · Quantitative Finance 2015-06-17 Shuhua Chang , Xinyu Wang , Alexander Shananin

In this paper we study a continuous time equilibrium model of limit order book (LOB) in which the liquidity dynamics follows a non-local, reflected mean-field stochastic differential equation (SDE) with evolving intensity. Generalizing the…

Optimization and Control · Mathematics 2020-03-03 Jin Ma , Eunjung Noh

We consider finite horizon Markov decision processes under performance measures that involve both the mean and the variance of the cumulative reward. We show that either randomized or history-based policies can improve performance. We prove…

Machine Learning · Computer Science 2011-05-02 Shie Mannor , John Tsitsiklis

We model the stock price dynamics through a semi-Markov process obtained using a Poisson random measure. We establish the existence and uniqueness of the classical solution of a non-homogeneous terminal value problem and we show that the…

Mathematical Finance · Quantitative Finance 2022-09-13 Garima Agrawal , Anindya Goswami

We propose a price impact model where changes in prices are purely driven by the order flow in the market. The stochastic price impact of market orders and the arrival rates of limit and market orders are functions of the market liquidity…

Trading and Market Microstructure · Quantitative Finance 2024-12-18 Peter Bank , Álvaro Cartea , Laura Körber

We study multi-objective reinforcement learning with nonlinear preferences over trajectories. That is, we maximize the expected value of a nonlinear function over accumulated rewards (expected scalarized return or ESR) in a multi-objective…

Machine Learning · Computer Science 2025-02-19 Nianli Peng , Muhang Tian , Brandon Fain

To sidestep the curse of dimensionality when computing solutions to Hamilton-Jacobi-Bellman partial differential equations (HJB PDE), we propose an algorithm that leverages a neural network to approximate the value function. We show that…

Machine Learning · Computer Science 2017-03-28 Frank Jiang , Glen Chou , Mo Chen , Claire J. Tomlin

Robust Markov decision processes (MDPs) address the challenge of model uncertainty by optimizing the worst-case performance over an uncertainty set of MDPs. In this paper, we focus on the robust average-reward MDPs under the model-free…

Machine Learning · Computer Science 2023-05-19 Yue Wang , Alvaro Velasquez , George Atia , Ashley Prater-Bennette , Shaofeng Zou

We study a class of backward stochastic differential equations (BSDEs) driven by a random measure or, equivalently, by a marked point process. Under appropriate assumptions we prove well-posedness and continuous dependence of the solution…

Probability · Mathematics 2012-05-24 Fulvia Confortola , Marco Fuhrman

We study value-iteration (VI) algorithms for solving general (a.k.a. multichain) Markov decision processes (MDPs) under the average-reward criterion, a fundamental but theoretically challenging setting. Beyond the difficulties inherent to…

Optimization and Control · Mathematics 2026-04-23 Matthew Zurek , Yudong Chen

We present a simple and easy to implement method for the numerical solution of a rather general class of Hamilton-Jacobi-Bellman (HJB) equations. In many cases, the considered problems have only a viscosity solution, to which, fortunately,…

Computational Finance · Quantitative Finance 2011-02-17 Jan Hendrik Witte , Christoph Reisinger

In this paper we study backward stochastic differential equations (BSDEs) driven by the compensated random measure associated to a given pure jump Markov process X on a general state space K. We apply these results to prove well-posedness…

Probability · Mathematics 2013-02-05 Fulvia Confortola , Marco Fuhrman

Discrete time stochastic optimal control problems and Markov decision processes (MDPs), respectively, serve as fundamental models for problems that involve sequential decision making under uncertainty and as such constitute the theoretical…

Optimization and Control · Mathematics 2023-03-08 Christian Beck , Arnulf Jentzen , Konrad Kleinberg , Thomas Kruse

Mean field game facilitates analyzing multi-armed bandit (MAB) for a large number of agents by approximating their interactions with an average effect. Existing mean field models for multi-agent MAB mostly assume a binary reward function,…

Multiagent Systems · Computer Science 2021-05-11 Xiong Wang , Riheng Jia

In reinforcement learning (RL), aligning agent behavior with specific objectives typically requires careful design of the reward function, which can be challenging when the desired objectives are complex. In this work, we propose an…

Machine Learning · Computer Science 2025-09-05 Yuting Tang , Yivan Zhang , Johannes Ackermann , Yu-Jie Zhang , Soichiro Nishimori , Masashi Sugiyama

A class of stochastic optimal control problems involving optimal stopping is considered. Methods of Krylov are adapted to investigate the numerical solutions of the corresponding normalized Bellman equations and to estimate the rate of…

Optimization and Control · Mathematics 2014-12-18 István Gyöngy , David Šiška

Motion planning under uncertainty for an autonomous system can be formulated as a Markov Decision Process with a continuous state space. In this paper, we propose a novel solution to this decision-theoretic planning problem that directly…

Robotics · Computer Science 2020-07-02 Junhong Xu , Kai Yin , Lantao Liu

The Ordered Upwind Method (OUM) is used to approximate the viscosity solution of the static Hamilton-Jacobi-Bellman (HJB) with direction-dependent weights on unstructured meshes. The method has been previously shown to provide a solution…

Optimization and Control · Mathematics 2016-01-13 Alex Shum , Kirsten Morris , Amir Khajepour

Controlling systems of ordinary differential equations (ODEs) is ubiquitous in science and engineering. For finding an optimal feedback controller, the value function and associated fundamental equations such as the Bellman equation and the…

Optimization and Control · Mathematics 2021-04-14 Mathias Oster , Leon Sallandt , Reinhold Schneider
‹ Prev 1 3 4 5 6 7 10 Next ›