English
Related papers

Related papers: Optimal Stopping with Rank-Dependent Loss

200 papers

We consider a class of reinforcement-learning systems in which the agent follows a behavior policy to explore a discrete state-action space to find an optimal policy while adhering to some restriction on its behavior. Such restriction may…

Machine Learning · Computer Science 2023-04-07 Peter C. Y. Chen

The problem of stopping a Brownian bridge with an unknown pinning point to maximise the expected value at the stopping time is studied. A few general properties, such as continuity and various bounds of the value function, are established.…

Probability · Mathematics 2019-01-17 Erik Ekström , Juozas Vaicenavicius

We consider adaptive decision-making problems where an agent optimizes a cumulative performance objective by repeatedly choosing among a finite set of options. Compared to the classical prediction-with-expert-advice set-up, we consider…

Machine Learning · Computer Science 2023-04-10 Michael Muehlebach

This paper considers an optimal impulse control problem of dynamical systems generated by a flow. The performance criteria are total costs over the infinite time horizon. Apart from the main performance to be minimized, there are multiple…

Optimization and Control · Mathematics 2020-10-27 Alexey Piunovskiy , Yi Zhang

Most clinical prediction studies are developed from retrospective cohorts and reported as if all patient information were observed at once. In practice, clinicians face a more consequential question: \emph{when is there already enough…

Methodology · Statistics 2026-04-27 Hui-Mean Foo , Yuan-chin Ivan Chang

The problem of statistical learning is to construct a predictor of a random variable $Y$ as a function of a related random variable $X$ on the basis of an i.i.d. training sample from the joint distribution of $(X,Y)$. Allowable predictors…

Information Theory · Computer Science 2016-11-15 Maxim Raginsky

In this paper we consider stopping problems with partial observation under a general risk-sensitive optimization criterion for problems with finite and infinite time horizon. Our aim is to maximize the certainty equivalent of the stopping…

Optimization and Control · Mathematics 2017-03-29 Nicole Bäuerle , Ulrich Rieder

In this article, a general problem of sequential statistical inference for general discrete-time stochastic processes is considered. The problem is to minimize an average sample number given that Bayesian risk due to incorrect decision does…

Statistics Theory · Mathematics 2010-10-18 Andrey Novikov

We study the optimal stopping problem for dynamic risk measures represented by Backward Stochastic Differential Equations (BSDEs) with jumps and its relation with reflected BSDEs (RBSDEs). We first provide general existence, uniqueness and…

Probability · Mathematics 2013-01-01 Marie-Claire Quenez , AgnÈs Sulem

The recent work by Dong & Yang (2023) showed for misspecified sparse linear bandits, one can obtain an $O\left(\epsilon\right)$-optimal policy using a polynomial number of samples when the sparsity is a constant, where $\epsilon$ is the…

Machine Learning · Computer Science 2024-07-19 Ally Yalei Du , Lin F. Yang , Ruosong Wang

Stop-loss rules are often studied in the financial literature, but the stop-loss levels are seldom constructed systematically. In many papers, and indeed in practice as well, the level of the stops is too often set arbitrarily. Guided by…

Risk Management · Quantitative Finance 2016-09-06 Antoine Emil Zambelli

It is well-known that Excess-of-Loss reinsurance has more marketability than Stop-Loss reinsurance, though Stop-Loss reinsurance is the most prominent setting discussed in the optimal (re)insurance design literature. We point out that…

Applications · Statistics 2024-05-02 Ernest Aboagye , Vali Asimit , Tsz Chai Fung , Liang Peng , Qiuqi Wang

We give a complete characterization of the complexity of best-arm identification in one-parameter bandit problems. We prove a new, tight lower bound on the sample complexity. We propose the `Track-and-Stop' strategy, which we prove to be…

Statistics Theory · Mathematics 2016-06-02 Aurélien Garivier , Emilie Kaufmann

Information-theoretic Bayesian regret bounds of Russo and Van Roy capture the dependence of regret on prior uncertainty. However, this dependence is through entropy, which can become arbitrarily large as the number of actions increases. We…

Machine Learning · Statistics 2020-07-09 Shi Dong , Benjamin Van Roy

In this paper we consider discrete and continuous time risk sensitive optimal stopping problem. Using suitable properties of the underlying Feller-Markov process we prove continuity of the optimal stopping value function and provide formula…

Optimization and Control · Mathematics 2021-03-31 Damian Jelito , Marcin Pitera , Łukasz Stettner

Many decision problems in economics, information technology, and industry can be transformed to an optimal stopping of adapted random vectors with some utility function over the set of Markov times with respect to filtration build by the…

Optimization and Control · Mathematics 2020-11-04 Krzysztof Szajowski

We extend the classical setting of an optimal stopping problem under full information to include for problems with an unknown state. The framework allows the unknown state to influence (i) the drift of the underlying process, (ii) the…

Probability · Mathematics 2024-05-08 Erik Ekström , Yuqiong Wang

An unconventional approach for optimal stopping under model ambiguity is introduced. Besides ambiguity itself, we take into account how ambiguity-averse an agent is. This inclusion of ambiguity attitude, via an $\alpha$-maxmin nonlinear…

Mathematical Finance · Quantitative Finance 2021-07-15 Yu-Jui Huang , Xiang Yu

This paper discusses the problem of determining optimal designs for regression models, when the observations are dependent and taken on an interval. A complete solution of this challenging optimal design problem is given for a broad class…

Methodology · Statistics 2015-02-25 Holger Dette , Andrey Pepelyshev , Anatoly Zhigljavsky

We consider an optimal stopping problem where a constraint is placed on the distribution of the stopping time. Reformulating the problem in terms of so-called measure-valued martingales allows us to transform the marginal constraint into an…

Optimization and Control · Mathematics 2017-03-27 Sigrid Källblad
‹ Prev 1 4 5 6 7 8 10 Next ›