中文
相关论文

相关论文: Optimal Stopping with Rank-Dependent Loss

200 篇论文

In 2019, Anderson et al. proposed the concept of rankability, which refers to a dataset's inherent ability to be meaningfully ranked. In this article, we give an expository review of the linear ordering problem (LOP) and then use it to…

最优化与控制 · 数学 2021-04-14 Thomas R. Cameron , Sebastian Charmot , Jonad Pulaj

We consider a nonlinear system, affine with respect to an unbounded control $u$ which is allowed to range in a closed cone. To this system we associate a Bolza type minimum problem, with a Lagrangian having sublinear growth with respect to…

最优化与控制 · 数学 2019-07-11 M. Soledad Aronna , Monica Motta , Franco Rampazzo

Given an initial (resp., terminal) probability measure $\mu$ (resp., $\nu$) on $\mathbb{R}^d$, we characterize those optimal stopping times $\tau$ that maximize or minimize the functional $\mathbb{E} |B_0 - B_\tau|^{\alpha}$, $\alpha > 0$,…

概率论 · 数学 2017-11-09 Nassif Ghoussoub , Young-Heon Kim , Tongseok Lim

Various algorithms in reinforcement learning exhibit dramatic variability in their convergence rates and ultimate accuracy as a function of the problem structure. Such instance-specific behavior is not captured by existing global minimax…

机器学习 · 统计学 2021-06-29 Koulik Khamaru , Eric Xia , Martin J. Wainwright , Michael I. Jordan

Let X_t, 0<=t<=T be a one-dimensional stochastic process with independent and stationary increments. This paper considers the problem of stopping the process X_t "as close as possible" to its eventual supremum M_T:=sup{X_t: 0<=t<=T}, when…

概率论 · 数学 2012-03-21 Pieter C. Allaart

We investigate an entropy-regularized reinforcement learning (RL) approach to optimal stopping problems motivated by real option models. Classical stopping rules are strict and non-randomized, limiting natural exploration in RL settings. To…

最优化与控制 · 数学 2026-02-18 Jodi Dianetti , Giorgio Ferrari , Renyuan Xu

This paper concerns an optimal stopping problem driven by the running maximum of a spectrally negative Levy process X. More precisely, we are interested in capped versions of the American lookback optimal stopping problem, which has its…

概率论 · 数学 2012-04-17 Andreas E. Kyprianou , Curdin Ott

Given a stable L\'{e}vy process $X=(X_t)_{0\le t\le T}$ of index $\alpha\in(1,2)$ with no negative jumps, and letting $S_t=\sup_{0\le s\le t}X_s$ denote its running supremum for $t\in [0,T]$, we consider the optimal prediction problem…

概率论 · 数学 2012-02-10 Violetta Bernyk , Robert C. Dalang , Goran Peskir

Given two probability measures $\mu, \nu$ on $\mathbb{R}^d$, in subharmonic order, we describe optimal stopping times $\tau$ that maximize/minimize the cost functional $\mathbb{E} |B_0 - B_\tau|^{\alpha}$, $\alpha > 0$, where $(B_t)_t$ is…

偏微分方程分析 · 数学 2019-06-28 Nassif Ghoussoub , Young-Heon Kim , Tongseok Lim

We propose and analyze a continuous-time robust reinforcement learning framework for optimal stopping under ambiguity. In this framework, an agent chooses a robust exploratory stopping time motivated by two objectives: robust…

最优化与控制 · 数学 2026-04-17 Junyan Ye , Hoi Ying Wong , Kyunghyun Park

We consider the optimal stopping problem for a Gauss-Markov process conditioned to adopt a prescribed terminal distribution. By applying a time-space transformation, we show it is equivalent to stopping a Brownian bridge pinned at a random…

概率论 · 数学 2025-05-26 Abel Azze , Bernardo D'Auria

In this paper we develop a deep learning method for optimal stopping problems which directly learns the optimal stopping rule from Monte Carlo samples. As such, it is broadly applicable in situations where the underlying randomness can…

数值分析 · 数学 2021-11-02 Sebastian Becker , Patrick Cheridito , Arnulf Jentzen

We consider the optimal stopping problem $v^{(\eps)}:=\sup_{\tau\in\mathcal{T}_{0,T}}\mathbb{E}B_{(\tau-\eps)^+}$ posed by Shiryaev at the International Conference on Advanced Stochastic Optimization Problems organized by the Steklov…

概率论 · 数学 2015-04-07 Erhan Bayraktar , Zhou Zhou

Suppose $N$ independent Bernoulli trials are observed sequentially at random times of a mixed binomial process. The task is to maximise, by using a nonanticipating stopping strategy, the probability of stopping at the last success. We focus…

概率论 · 数学 2024-10-22 Alexander Gnedin , Zakaria Derbazi

For an infinite-horizon continuous-time optimal stopping problem under non-exponential discounting, we look for an optimal equilibrium, which generates larger values than any other equilibrium does on the entire state space. When the…

最优化与控制 · 数学 2021-07-15 Yu-Jui Huang , Zhou Zhou

In real-world reinforcement learning (RL) systems, various forms of {\it impaired observability} can complicate matters. These situations arise when an agent is unable to observe the most recent state of the system due to latency or lossy…

机器学习 · 计算机科学 2023-10-30 Minshuo Chen , Jie Meng , Yu Bai , Yinyu Ye , H. Vincent Poor , Mengdi Wang

We consider the L\'evy model of the perpetual American call and put options with a negative discount rate under Poisson observations. Similar to the continuous observation case as in De Donno et al. [24], the stopping region that…

最优化与控制 · 数学 2020-04-08 Zbigniew Palmowski , José Luis Pérez , Kazutoshi Yamazaki

We solve an optimal stopping problem where the underlying diffusion is Brownian motion on $\bf R$ with a positive drift changing at zero. It is assumed that the drift $\mu_1$ on the negative side is smaller than the drift $\mu_2$ on the…

概率论 · 数学 2018-11-15 Ernesto Mordecki , Paavo Salminen

We introduce a notion of bounded variation solution for a new class of nonlinear control systems with ordinary and impulsive controls, in which the drift function depends not only on the state, but also on its past history, through a finite…

最优化与控制 · 数学 2023-07-25 Giovanni Fusco , Monica Motta

We consider the problem of binary classification with abstention in the relatively less studied \emph{bounded-rate} setting. We begin by obtaining a characterization of the Bayes optimal classifier for an arbitrary input-label distribution…

机器学习 · 计算机科学 2019-05-24 Shubhanshu Shekhar , Mohammad Ghavamzadeh , Tara Javidi