中文
相关论文

相关论文: A variational formula for risk-sensitive reward

200 篇论文

We consider a problem of optimal control of an infinite horizon system governed by forward-backward stochastic differential equations with delay. Sufficient and necessary maximum principles for optimal control under partial information in…

最优化与控制 · 数学 2013-12-09 Nacira Agram , Bernt Øksendal

In this article, we consider the deterministic impulsively controlled system with infinite horizon and several discounted objective functionals. The constructed optimal control problem with functional constraints is reformulated as a Markov…

最优化与控制 · 数学 2026-02-10 A. Piunovskiy

We consider a long-run impulse control problem for a generic Markov process with a multiplicative reward functional. We construct a solution to the associated Bellman equation and provide a verification result. The argument is based on the…

最优化与控制 · 数学 2023-05-15 Damian Jelito , Łukasz Stettner

For a semi-martingale $X_t$, which forms a stochastic boundary, a rate-optimal estimator for its quadratic variation $\langle X, X \rangle_t$ is constructed based on observations in the vicinity of $X_t$. The problem is embedded in a…

概率论 · 数学 2015-11-24 Markus Bibinger , Moritz Jirak , Markus Reiß

This paper studies a risk-sensitive decision-making problem under uncertainty. It considers a decision-making process that unfolds over a fixed number of stages, in which a decision-maker chooses among multiple alternatives, some of which…

最优化与控制 · 数学 2026-01-07 Chung-Han Hsieh , Yi-Shan Wong

This paper studies an optimal dividend problem with a drawdown constraint in a Brownian motion model, requiring the dividend payout rate to remain above a fixed proportion of its historical maximum. This leads to a path-dependent stochastic…

数理金融 · 定量金融 2026-01-08 Chonghu Guan , Jiacheng Fan , Zuo Quan Xu

In this manuscript we consider a class optimal control problem for stochastic differential delay equations. First, we rewrite the problem in a suitable infinite-dimensional Hilbert space. Then, using the dynamic programming approach, we…

最优化与控制 · 数学 2023-02-20 Filippo de Feo , Salvatore Federico , Andrzej Święch

We consider a two-sided singular stochastic control problem with a risk-sensitive ergodic criterion. In particular, we consider a stochastic system whose uncontrolled dynamics are modelled by a linear diffusion. The control that can be…

最优化与控制 · 数学 2025-09-15 Justin Gwee , Mihail Zervos

We build optimal exponential bounds for the probabilities of large deviations of sums \sum_{k=1}^nf(X_k) where (X_k) is a finite reversible Markov chain and f is an arbitrary bounded function. These bounds depend only on the stationary mean…

概率论 · 数学 2007-05-23 Carlos A. Leon , Francois Perron

In this paper, we obtain the maximum principle for optimal controls of stochastic systems with jumps by introducing a new method of variation. The control is allowed to enter both diffusion and jump term and the control domain need not to…

最优化与控制 · 数学 2019-10-10 Yuanzhuo Song , Shanjian Tang , Zhen Wu

We introduce a general framework for measuring risk in the context of Markov control processes with risk maps on general Borel spaces that generalize known concepts of risk measures in mathematical finance, operations research and…

最优化与控制 · 数学 2014-01-27 Yun Shen , Wilhelm Stannat , Klaus Obermayer

This paper shows the usefulness of the Perov contraction theorem, which is a generalization of the classical Banach contraction theorem, for solving Markov dynamic programming problems. When the reward function is unbounded, combining an…

最优化与控制 · 数学 2024-05-06 Alexis Akira Toda

In the Markov decision process model, policies are usually evaluated by expected cumulative rewards. As this decision criterion is not always suitable, we propose in this paper an algorithm for computing a policy optimal for the quantile…

人工智能 · 计算机科学 2016-12-02 Hugo Gilbert , Paul Weng , Yan Xu

In this work, we address risk-averse Bayes-adaptive reinforcement learning. We pose the problem of optimising the conditional value at risk (CVaR) of the total return in Bayes-adaptive Markov decision processes (MDPs). We show that a policy…

机器学习 · 计算机科学 2021-10-27 Marc Rigter , Bruno Lacerda , Nick Hawes

We consider a general class of dynamic resource allocation problems within a stochastic optimal control framework. This class of problems arises in a wide variety of applications, each of which intrinsically involves resources of different…

最优化与控制 · 数学 2018-01-08 Xuefeng Gao , Yingdong Lu , Mayank Sharma , Mark S. Squillante , Joost W. Bosman

Linear dynamical systems that obey stochastic differential equations are canonical models. While optimal control of known systems has a rich literature, the problem is technically hard under model uncertainty and there are hardly any…

系统与控制 · 电气工程与系统科学 2023-06-09 Mohamad Kazem Shirani Faradonbeh , Mohamad Sadegh Shirani Faradonbeh

We propose a new model for formalizing reward collection problems on graphs with dynamically generated rewards which may appear and disappear based on a stochastic model. The *robot routing problem* is modeled as a graph whose nodes are…

系统与控制 · 计算机科学 2017-07-18 Rayna Dimitrova , Ivan Gavran , Rupak Majumdar , Vinayak S. Prabhu , Sadegh Esmaeil Zadeh Soudjani

This paper addresses the problem of utility maximization under uncertain parameters. In contrast with the classical approach, where the parameters of the model evolve freely within a given range, we constrain them via a penalty function. We…

最优化与控制 · 数学 2022-03-08 Ivan Guo , Nicolas Langrené , Grégoire Loeper , Wei Ning

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

最优化与控制 · 数学 2025-12-18 Lingjia Zeng , Manman Li

Decision-theoretic planning with risk-sensitive planning objectives is important for building autonomous agents or decision-support systems for real-world applications. However, this line of research has been largely ignored in the…

人工智能 · 计算机科学 2012-07-09 Yaxin Liu , Sven Koenig
‹ 上一页 1 8 9 10 下一页 ›