English
Related papers

Related papers: A new strategy for Robbins' problem of optimal sto…

200 papers

The empirical loss, commonly referred to as the average loss, is extensively utilized for training machine learning models. However, in order to address the diverse performance requirements of machine learning models, the use of the…

Optimization and Control · Mathematics 2024-01-04 Rufeng Xiao , Yuze Ge , Rujun Jiang , Yifan Yan

We investigate an entropy-regularized reinforcement learning (RL) approach to optimal stopping problems motivated by real option models. Classical stopping rules are strict and non-randomized, limiting natural exploration in RL settings. To…

Optimization and Control · Mathematics 2026-02-18 Jodi Dianetti , Giorgio Ferrari , Renyuan Xu

We present a novel method for solving a class of time-inconsistent optimal stopping problems by reducing them to a family of standard stochastic optimal control problems. In particular, we convert an optimal stopping problem with a…

Optimization and Control · Mathematics 2016-11-15 Christopher W. Miller

Many decision problems in economics, information technology, and industry can be transformed to an optimal stopping of adapted random vectors with some utility function over the set of Markov times with respect to filtration build by the…

Optimization and Control · Mathematics 2020-11-04 Krzysztof Szajowski

A version of the classical secretary problem is studied, in which one is interested in selecting one of the b best out of a group of n differently ranked persons who are presented one by one in a random order. It is assumed that b is a…

Probability · Mathematics 2010-09-06 Chris Dietz , Dinard van der Laan , Ad Ridder

The importance of a node in a directed graph can be measured by its PageRank. The PageRank of a node is used in a number of application contexts - including ranking websites - and can be interpreted as the average portion of time spent at…

Data Structures and Algorithms · Computer Science 2014-05-22 Balázs Csanád Csáji , Raphaël M. Jungers , Vincent D. Blondel

We analyze an optimal stopping problem with random maturity under a nonlinear expectation with respect to a weakly compact set of mutually singular probabilities $\mathcal{P}$. The maturity is specified as the hitting time to level $0$ of…

Probability · Mathematics 2016-07-08 Erhan Bayraktar , Song Yao

Graph planning gives rise to fundamental algorithmic questions such as shortest path, traveling salesman problem, etc. A classical problem in discrete planning is to consider a weighted graph and construct a path that maximizes the sum of…

Artificial Intelligence · Computer Science 2018-02-13 Krishnendu Chatterjee , Laurent Doyen

In the first part of this paper, we study RBSDEs in the case where the filtration is not quasi-left continuous and the lower obstacle is given by a predictable process. We prove the existence and uniqueness by using some results of optimal…

Probability · Mathematics 2018-12-03 S. Bouhadou , Y. Ouknine

In this article we study forbidden loci and typical ranks of forms with respect to the embeddings of $\mathbb P^1\times \mathbb P^1$ given by the line bundles $(2,2d)$. We introduce the Ranestad-Schreyer locus corresponding to supports of…

Algebraic Geometry · Mathematics 2018-03-16 Emanuele Ventura

This paper considers the use of recently proposed optimal transport-based multivariate test statistics, namely rank energy and its variant the soft rank energy derived from entropically regularized optimal transport, for the unsupervised…

Machine Learning · Statistics 2023-02-17 Matthew Werenski , Shoaib Bin Masud , James M. Murphy , Shuchin Aeron

Suppose $N$ independent Bernoulli trials are observed sequentially at random times of a mixed binomial process. The task is to maximise, by using a nonanticipating stopping strategy, the probability of stopping at the last success. We focus…

Probability · Mathematics 2024-10-22 Alexander Gnedin , Zakaria Derbazi

The classical secretary problem has been generalized over the years into several directions. In this paper we confine our interest to those generalizations which have to do with the more general problem of stopping on a last observation of…

Performance · Computer Science 2017-05-29 Guy Louchard

In this paper we introduce and solve a class of optimal stopping problems of recursive type. In particular, the stopping payoff depends directly on the value function of the problem itself. In a multi-dimensional Markovian setting we show…

Optimization and Control · Mathematics 2021-06-23 Katia Colaneri , Tiziano De Angelis

We consider optimal stopping problems, in which a sequence of independent random variables is drawn from a known continuous density. The objective of such problems is to find a procedure which maximizes the expected reward; this is often…

Probability · Mathematics 2020-12-07 Hugh Entwistle , Christopher Lustri , Georgy Sofronov

In the best choice problem with random arrivals, an unknown number $n$ of rankable items arrive at times sampled from the uniform distribution. As is well known, a real-time player can ensure stopping at the overall best item with…

Probability · Mathematics 2021-03-09 Alexander Gnedin

We consider the problem of recovering an unknown matching between a set of $n$ randomly placed points in $\mathbb{R}^d$ and random perturbations of these points. This can be seen as a model for particle tracking and more generally, entity…

Statistics Theory · Mathematics 2024-03-27 Lucas da Rocha Schwengber , Roberto Imbuzeiro Oliveira

We extend the ideas of Diening, Kreuzer, and Stevenson [Instance optimality of the adaptive maximum strategy, Found. Comput. Math. (2015)], from conforming approximations of the Poisson problem to nonconforming Crouzeix-Raviart…

Numerical Analysis · Mathematics 2015-04-13 Christian Kreuzer , Mira Schedensack

In this paper, we propose how to use objective arguments grounded in statistical mechanics concepts in order to obtain a single number, obtained after aggregation, which would allow to rank "agents", "opinions", ..., all defined in a very…

Physics and Society · Physics 2024-05-02 Marcel Ausloos , Giulia Rotundo , Roy Cerqueti

We develop a decision making framework to cast the problem of learning a ranking policy for search or recommendation engines in a two-sided e-commerce marketplace as an expected reward optimization problem using observational data. As a…

Information Retrieval · Computer Science 2024-10-08 Ehsan Ebrahimzadeh , Nikhil Monga , Hang Gao , Alex Cozzi , Abraham Bagherjeiran
‹ Prev 1 4 5 6 7 8 10 Next ›