English
Related papers

Related papers: Long-Run Average Reward Maximization of A Regulate…

200 papers

In this paper we study a class of optimal dividend and investment problems assuming that the underlying reserve process follows the Sparre Andersen model, that is, the claim frequency is a "renewal" process, rather than a standard compound…

Probability · Mathematics 2016-07-05 Lihua Bai , Jin Ma , Xiaojing Xing

This paper studies a type of rank-based mean field game in which competing agents strategically switch among multiple effort regimes. We propose an entropy regularized auxiliary problem where the switching decisions are randomized to the…

Optimization and Control · Mathematics 2026-05-29 Zongxia Liang , Shu Wang , Xiang Yu

We develop an approach for two player constraint zero-sum and nonzero-sum stochastic differential games, which are modeled by Markov regime-switching jump-diffusion processes. We provide the relations between a usual stochastic optimal…

Optimization and Control · Mathematics 2023-01-31 Emel Savku

We present a case study applying learning-based distributionally robust model predictive control to highway motion planning under stochastic uncertainty of the lane change behavior of surrounding road users. The dynamics of road users are…

Systems and Control · Electrical Eng. & Systems 2022-11-08 Mathijs Schuurmans , Alexander Katriniok , Christopher Meissen , H. Eric Tseng , Panagiotis Patrinos

This contribution mainly focuses on the finite horizon optimal control problems of a susceptible-infected-vaccinated(SIV) epidemic system governed by reaction-diffusion equations and Markov switching. Stochastic dynamic programming is…

Optimization and Control · Mathematics 2024-01-23 Zong Wang

In this work, we consider the optimal portfolio selection problem under hard constraints on trading volume amounts when the dynamics of the risky asset returns are governed by a discrete-time approximation of the Markov-modulated geometric…

Portfolio Management · Quantitative Finance 2014-10-07 Vladimir Dombrovskii , Tatyana Obyedko

Stochastic optimal control problems have a long tradition in applied probability, with the questions addressed being of high relevance in a multitude of fields. Even though theoretical solutions are well understood in many scenarios, their…

Statistics Theory · Mathematics 2024-05-28 Sören Christensen , Claudia Strauch , Lukas Trottner

Option-critic learning is a general-purpose reinforcement learning (RL) framework that aims to address the issue of long term credit assignment by leveraging temporal abstractions. However, when dealing with extended timescales, discounting…

Machine Learning · Computer Science 2019-11-21 Akshay Dharmavaram , Matthew Riemer , Shalabh Bhatnagar

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

We consider a portfolio optimization problem in a defaultable market with finitely-many economical regimes, where the investor can dynamically allocate her wealth among a defaultable bond, a stock, and a money market account. The market…

Portfolio Management · Quantitative Finance 2011-09-07 Agostino Capponi , Jose E. Figueroa-Lopez

We consider the problem of controlling a Markov decision process (MDP) with a large state space, so as to minimize average cost. Since it is intractable to compete with the optimal policy for large scale problems, we pursue the more modest…

Optimization and Control · Mathematics 2014-02-28 Yasin Abbasi-Yadkori , Peter L. Bartlett , Alan Malek

This paper deals with discrete-time Markov control processes on a general state space. A long-run risk-sensitive average cost criterion is used as a performance measure. The one-step cost function is nonnegative and possibly unbounded.…

Risk Management · Quantitative Finance 2016-08-14 Anna Jaśkiewicz

The optimization criterion for dividends from a risky business is most often formalized in terms of the expected present value of future dividends. That criterion disregards a potential, explicit demand for stability of dividends. In…

Optimization and Control · Mathematics 2023-06-22 Benjamin Avanzi , Debbie Kusch Falden , Mogens Steffensen

This paper deals with numerical solutions of maximizing expected utility from terminal wealth under a non-bankruptcy constraint. The wealth process is subject to shocks produced by a general marked point process. The problem of the agent is…

Computational Finance · Quantitative Finance 2010-09-06 Mohamed Mnif

We consider a singular control problem with regime switching that arises in problems of optimal investment decisions of cash-constrained firms. The value function is proved to be the unique viscosity solution of the associated…

Computational Finance · Quantitative Finance 2016-10-07 Erwan Pierre , Stéphane Villeneuve , Xavier Warin

We introduce the Lyapunov approach to optimal control problems of average risk-sensitive Markov control processes with general risk maps. Motivated by applications in particular to behavioral economics, we consider possibly non-convex risk…

Optimization and Control · Mathematics 2015-07-23 Yun Shen , Klaus Obermayer , Wilhelm Stannat

This paper addresses objectives tailored to the risk-averse optimization of accumulated rewards in Markov decision processes (MDPs). The studied objectives require maximizing the expected value of the accumulated rewards minus a penalty…

Logic in Computer Science · Computer Science 2024-07-10 Christel Baier , Jakob Piribauer , Maximilian Starke

This paper analyzes and explicitly solves a class of long-term average impulse control problems and a related class of singular control problems. The underlying process is a general one-dimensional diffusion with appropriate boundary…

Optimization and Control · Mathematics 2026-05-05 K. L. Helmes , R. H. Stockbridge , C. Zhu

This paper considers an optimal control of a big financial company with debt liability under bankrupt probability constraints. The company, which faces constant liability payments and has choices to choose various production/business…

Risk Management · Quantitative Finance 2010-08-11 Zongxia Liang , Bin Sun

We study a an optimal high frequency trading problem within a market microstructure model designed to be a good compromise between accuracy and tractability. The stock price is driven by a Markov Renewal Process (MRP), while market orders…

Trading and Market Microstructure · Quantitative Finance 2015-01-06 Pietro Fodra , Huyên Pham