English
Related papers

Related papers: Constrained Optimal Stopping, Liquidity and Effort

200 papers

In this paper we consider stochastic optimization problems for an ambiguity averse decision maker who is uncertain about the parameters of the underlying process. In a first part we consider problems of optimal stopping under drift…

Computational Finance · Quantitative Finance 2015-03-19 Sören Christensen

We introduce a novel extension of the canonical multi-armed bandit problem that incorporates an additional strategic innovation: abstention. In this enhanced framework, the agent is not only tasked with selecting an arm at each time step,…

Machine Learning · Computer Science 2026-03-24 Junwen Yang , Tianyuan Jin , Vincent Y. F. Tan

We study the properties of the free boundaries and the corresponding hitting times in the context of optimal stopping in discrete time. We first prove the continuity of the map from the boundaries to the expected value of the corresponding…

Probability · Mathematics 2025-04-16 H. Mete Soner , Valentin Tissot-Daguette

We describe the solution of an optimal stopping problem for a stable L\'evy process killed at state-dependent rate, which can be seen as a model for bankruptcy. The killing rate is chosen in such a way that the killed process remains…

Probability · Mathematics 2024-02-29 K. van Schaik , A. R. Watson , X. Xu

We consider the problem of designing a sequential decision making agent to maximize an unknown time-varying function which switches with time. At each step, the agent receives an observation of the function's value at a point decided by the…

Optimization and Control · Mathematics 2023-11-07 Durgesh Kalwar , Vineeth B. S

A random walk (or a Wiener process), possibly with drift, is observed in a noisy or delayed fashion. The problem considered in this paper is to estimate the first time \tau the random walk reaches a given level. Specifically, the p-moment…

Information Theory · Computer Science 2012-03-22 Marat V. Burnashev , Aslan Tchamkerten

We study a linear price impact model including other liquidity takers, whose flow of orders either follows a Poisson or a Hawkes process. The optimal execution problem is solved explicitly in this context, and the closed-formula optimal…

Trading and Market Microstructure · Quantitative Finance 2015-06-10 Aurélien Alfonsi , Pierre Blanc

We develop the first Bayesian Optimization algorithm, BLOSSOM, which selects between multiple alternative acquisition functions and traditional local optimization at each step. This is combined with a novel stopping condition based on…

Machine Learning · Statistics 2018-05-23 Mark McLeod , Michael A. Osborne , Stephen J. Roberts

We introduce a simple stochastic volatility model, whose novelty consists in taking into account hitting times of the asset price, and study the optimal stopping problem corresponding to a put option whose time horizon (after the asset…

Pricing of Securities · Quantitative Finance 2017-03-29 Sigurd Assing , Yufan Zhao

Sample size determination is crucial in experimental design, especially in traffic and transport research. Frequentist statistics require a fixed sample size determined by power analysis, which cannot be adjusted once the experiment starts.…

Methodology · Statistics 2025-03-04 Xiaomi Yang , Carol Flannagan , Jonas Bärgman

We study the optimal portfolio liquidation problem over a finite horizon in a limit order book with bid-ask spread and temporary market price impact penalizing speedy execution trades. We use a continuous-time modeling framework, but in…

Probability · Mathematics 2014-01-10 Idris Kharroubi , Huyen Pham

Originally motivated by default risk management applications, this paper investigates a novel problem, referred to as the profitable bandit problem here. At each step, an agent chooses a subset of the K possible actions. For each action…

Machine Learning · Statistics 2018-05-09 Mastane Achab , Stephan Clémençon , Aurélien Garivier

We formulate an optimal stopping problem for a geometric Brownian motion where the probability scale is distorted by a general nonlinear function. The problem is inherently time inconsistent due to the Choquet integration involved. We…

Probability · Mathematics 2022-01-07 Zuo Quan Xu , Xun Yu Zhou

Agents that learn to select optimal actions represent a prominent focus of the sequential decision-making literature. In the face of a complex environment or constraints on time and resources, however, aiming to synthesize such an optimal…

Machine Learning · Computer Science 2021-06-23 Dilip Arumugam , Benjamin Van Roy

We consider a version of the continuum armed bandit where an action induces a filtered realisation of a non-homogeneous Poisson process. Point data in the filtered sample are then revealed to the decision-maker, whose reward is the total…

Machine Learning · Computer Science 2020-07-21 James A. Grant , Roberto Szechtman

This note considers a variation of the full-information secretary problem where the random variables to be observed are independent and identically distributed. Consider $X_1,\dots,X_n$ to be an independent sequence of random variables, let…

Probability · Mathematics 2017-09-11 José A. Islas

In classic reinforcement learning algorithms, agents make decisions at discrete and fixed time intervals. The duration between decisions becomes a crucial hyperparameter, as setting it too short may increase the problem's difficulty by…

Machine Learning · Computer Science 2023-10-26 Amirmohammad Karimi , Jun Jin , Jun Luo , A. Rupam Mahmood , Martin Jagersand , Samuele Tosatto

We introduce the problem of assigning resources to improve their utilization. The motivation comes from settings where agents have uncertainty about their own values for using a resource, and where it is in the interest of a group that…

Computer Science and Game Theory · Computer Science 2018-11-02 Hongyao Ma , Reshef Meir , David C. Parkes , James Zou

This paper is devoted to studying constrained continuous-time Markov decision processes (MDPs) in the class of randomized policies depending on state histories. The transition rates may be unbounded, the reward and costs are admitted to be…

Probability · Mathematics 2012-01-04 Xianping Guo , Xinyuan Song

In distributed model predictive control (DMPC), where a centralized optimization problem is solved in distributed fashion using dual decomposition, it is important to keep the number of iterations in the solution algorithm, i.e. the amount…

Optimization and Control · Mathematics 2013-07-11 Pontus Giselsson , Anders Rantzer