English
Related papers

Related papers: An analytic recursive method for optimal multiple …

200 papers

We study the dual model with capital injection under the additional condition that the dividend strategy is absolutely continuous. We consider a refraction-reflection strategy that pays dividends at the maximal rate whenever the surplus is…

Optimization and Control · Mathematics 2016-08-24 José-Luis Pérez , Kazutoshi Yamazaki

In this article, we study the classical finite-horizon optimal stopping problem for multidimensional diffusions through an approach that differs from what is typically found in the literature. More specifically, we first prove a key…

Optimization and Control · Mathematics 2025-03-05 Andrea Cosso , Laura Perelli

De Finetti's optimal dividend problem has recently been extended to the case dividend payments can only be made at Poisson arrival times. This paper considers the version with bail-outs where the surplus must be nonnegative uniformly in…

Probability · Mathematics 2018-01-03 Kei Noba , José-Luis Pérez , Kazutoshi Yamazaki , Kouji Yano

Bayesian curve fitting plays an important role in inverse problems, and is often addressed using the Reversible Jump Markov Chain Monte Carlo (RJMCMC) algorithm. However, this algorithm can be computationally inefficient without…

Computation · Statistics 2024-02-28 Zhiyao Tian , Anthony Lee , Shunhua Zhou

Explicit solution of an infinite horizon optimal stopping problem for a Levy processes with a polynomial reward function is given, in terms of the overall supremum of the process, when the solution of the problem is one-sided. The results…

Probability · Mathematics 2015-07-23 Ernesto Mordecki , Yuliya Mishura

In this paper, we derive a practical, general framework for creating adaptive iterative (linearization or splitting) algorithms to solve multi-physics problems. This means that, given an iterative method, we derive \textit{a posteriori}…

Numerical Analysis · Mathematics 2026-01-26 Jakob S. Stokke , Kundan Kumar , Florin A. Radu

The objective in this paper is to obtain fast converging reinforcement learning algorithms to approximate solutions to the problem of discounted cost optimal stopping in an irreducible, uniformly ergodic Markov chain, evolving on a compact…

Systems and Control · Computer Science 2019-10-01 Shuhang Chen , Adithya M. Devraj , Ana Bušić , Sean P. Meyn

Recursive retraining of generative models poses a critical representation challenge: when synthetic outputs are curated based on a fixed reward signal, the model tends to collapse onto a narrow set of outputs that over-optimize that…

Machine Learning · Computer Science 2026-05-11 Ali Falahati , Mohammad Mohammadi Amiri , Kate Larson , Lukasz Golab

We study statistical model checking of continuous-time stochastic hybrid systems. The challenge in applying statistical model checking to these systems is that one cannot simulate such systems exactly. We employ the multilevel Monte Carlo…

Systems and Control · Computer Science 2017-06-27 Sadegh Esmaeil Zadeh Soudjani , Rupak Majumdar , Tigran Nagapetyan

This paper develops an approach for solving perpetual discounted optimal stopping problems for multidimensional diffusions, with special emphasis on the $d$-dimensional Wiener process. We first obtain some verification theorems for…

Probability · Mathematics 2016-11-04 Sören Christensen , Fabián Crocce , Ernesto Mordecki , Paavo Salminen

This paper considers online optimization for a system that performs a sequence of back-to-back tasks. Each task can be processed in one of multiple processing modes that affect the duration of the task, the reward earned, and an additional…

Optimization and Control · Mathematics 2024-01-17 Michael J. Neely

The treatment assignment mechanism in a randomized clinical trial can be optimized for statistical efficiency within a specified class of randomization mechanisms. Optimal designs of this type have been characterized in terms of the…

Methodology · Statistics 2025-09-03 Wei Zhang , Zhiwei Zhang , Aiyi Liu

Solving optimal control problems to determine a stabilizing controller involves a significant computational effort. Time-varying optimal control provides a remedy by designing a tracking system, given as an ordinary differential equation,…

Systems and Control · Electrical Eng. & Systems 2026-04-16 Patrick Schmidt , Stefan Streif

We introduce a new method to price American-style options on underlying investments governed by stochastic volatility (SV) models. The method does not require the volatility process to be observed. Instead, it exploits the fact that the…

Computational Finance · Quantitative Finance 2012-07-26 Bhojnarine R. Rambharat , Anthony E. Brockwell

One major limitation to the applicability of Reinforcement Learning (RL) to many practical domains is the large number of samples required to learn an optimal policy. To address this problem and improve learning efficiency, we consider a…

Machine Learning · Computer Science 2023-08-07 Roberto Cipollone , Giuseppe De Giacomo , Marco Favorito , Luca Iocchi , Fabio Patrizi

Reinforcement Learning algorithms that learn from human feedback (RLHF) need to be efficient in terms of statistical complexity, computational complexity, and query complexity. In this work, we consider the RLHF setting where the feedback…

Machine Learning · Computer Science 2024-03-14 Runzhe Wu , Wen Sun

We revisit the problem of sampling from a target distribution that has a smooth strongly log-concave density everywhere in $\mathbb R^p$. In this context, if no additional density information is available, the randomized midpoint…

Statistics Theory · Mathematics 2023-06-19 Lu Yu , Avetik Karagulyan , Arnak Dalalyan

We study the problem of optimal state-feedback tracking control for unknown discrete-time deterministic systems with input constraints. To handle input constraints, state-of-art methods utilize a certain nonquadratic stage cost function,…

Systems and Control · Electrical Eng. & Systems 2020-12-09 Alexandros Tanzanakis , John Lygeros

In this work we consider optimal stopping problems with conditional convex risk measures called optimised certainty equivalents. Without assuming any kind of time-consistency for the underlying family of risk measures, we derive a novel…

Mathematical Finance · Quantitative Finance 2014-12-16 Denis Belomestny , Volker Kraetschmer

In this article we consider a toy example of an optimal stopping problem driven by fragmentation processes. We show that one can work with the concept of stopping lines to formulate the notion of an optimal stopping problem and moreover, to…

Probability · Mathematics 2011-01-27 Andreas E. Kyprianou , Juan Carlos Pardo