English
Related papers

Related papers: Grab It Before It's Gone: Testing Uncertain Reward…

200 papers

We consider the problem of optimally stopping a general one-dimensional stochastic differential equation (SDE) with generalised drift over an infinite time horizon. First, we derive a complete characterisation of the solution to this…

Probability · Mathematics 2019-09-26 Mihail Zervos , Neofytos Rodosthenous , Pui Chan Lon , Thomas Bernhardt

We develop a probabilistic framework for analysing model-based reinforcement learning in the episodic setting. We then apply it to study finite-time horizon stochastic control problems with linear dynamics but unknown coefficients and…

Machine Learning · Computer Science 2021-12-22 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

This paper presents a new performance bound for estimation problems where the parameter to estimate lies in a Riemannian manifold (a smooth manifold endowed with a Riemannian metric) and follows a given prior distribution. In this setup,…

Statistics Theory · Mathematics 2024-09-10 Florent Bouchard , Alexandre Renaux , Guillaume Ginolhac , Arnaud Breloy

Price determination is a central research topic of revenue management in marketing. The important aspect in pricing is controlling the stochastic behavior of demand, and the previous studies have tackled price optimization problems with…

Optimization and Control · Mathematics 2024-01-04 Yuya Hikima , Akiko Takeda

We consider one-dimensional stochastic differential equations with a boundary condition, driven by a Poisson process. We study existence and uniqueness of solutions and the absolute continuity of the law of the solution. In the case when…

Probability · Mathematics 2007-05-23 Aureli Alabert , Miguel A. Marmolejo

If a variational problem comes with no boundary conditions prescribed beforehand, and yet these arise as a consequence of the variation process itself, we speak of a free boundary values variational problem. Such is, for instance, the…

Differential Geometry · Mathematics 2017-03-14 Giovanni Moreno , Monika Ewa Stypa

Experimental design is crucial for inference where limitations in the data collection procedure are present due to cost or other restrictions. Optimal experimental designs determine parameters that in some appropriate sense make the data…

Machine Learning · Statistics 2016-03-11 Panagiotis Tsilifis , Roger G. Ghanem , Paris Hajali

We solve two stochastic control problems in which a player tries to minimize or maximize the exit time from an interval of a Brownian particle, by controlling its drift. The player can change from one drift to another but is subject to a…

Probability · Mathematics 2014-08-19 Robert C. Dalang , Laura Vinckenbosch

We consider an individual or household endowed with an initial capital and an income, modeled as a deterministic process with a continuous drift rate. At first, we model the discounting rate as the price of a zero-coupon bond at zero under…

Optimization and Control · Mathematics 2016-04-01 Julia Eisenberg

We will study a free boundary value problem driven by a source term which is quite {\it irregular}. In the process, we will establish a monotonicity result, and regularity of the solution.

Analysis of PDEs · Mathematics 2024-04-15 Debajyoti Choudhuri , Shengda Zeng

We consider stochastic optimization under distributional uncertainty, where the unknown distributional parameter is estimated from streaming data that arrive sequentially over time. Moreover, data may depend on the decision of the time when…

Optimization and Control · Mathematics 2023-10-17 Tianyi Liu , Yifan Lin , Enlu Zhou

Stochastic dominance serves as a general framework for modeling a broad spectrum of decision preferences under uncertainty, with risk aversion as one notable example, as it naturally captures the intrinsic structure of the underlying…

Machine Learning · Computer Science 2026-01-06 Shicong Cen , Jincheng Mei , Hanjun Dai , Dale Schuurmans , Yuejie Chi , Bo Dai

Signal processing makes extensive use of point estimators and accompanying error bounds. These work well up until the likelihood function has two or more high peaks. When it is important for an estimator to remain reliable, it becomes…

Methodology · Statistics 2025-03-04 Ning Xu , Christopher M. Foster , Jonathan H. Manton

We consider the dynamic linear regression problem, where the predictor vector may vary with time. This problem can be modeled as a linear dynamical system, with non-constant observation operator, where the parameters that need to be learned…

Machine Learning · Computer Science 2022-10-13 Mark Kozdoba , Edward Moroshko , Shie Mannor , Koby Crammer

The sporadic task model is often used to analyze recurrent execution of identical tasks in real-time systems. A sporadic task defines an infinite sequence of task instances, also called jobs, that arrive under the minimum inter-arrival time…

Operating Systems · Computer Science 2018-03-01 Jian-Jia Chen , Georg von der Brüggen , Niklas Ueter

We study a Weiner process that is conditioned to pass through a finite set of points and consider the dynamics generated by iterating a sample path from this process. Using topological techniques we are able to characterize the global…

Dynamical Systems · Mathematics 2022-09-26 Konstantin Mischaikow , Cameron Thieme

Reward design is a critical part of the application of reinforcement learning, the performance of which strongly depends on how well the reward signal frames the goal of the designer and how well the signal assesses progress in reaching…

Machine Learning · Computer Science 2022-08-01 Yixiang Wang , Yujing Hu , Feng Wu , Yingfeng Chen

An important step in the Markov reward approach to error bounds on stationary performance measures of Markov chains is to bound the bias terms. Affine functions have been successfully used for these bounds for various models, but there are…

Probability · Mathematics 2019-01-04 Xinwei Bai , Jasper Goseling

The window mechanism was introduced by Chatterjee et al. to strengthen classical game objectives with time bounds. It permits to synthesize system controllers that exhibit acceptable behaviors within a configurable time frame, all along…

Logic in Computer Science · Computer Science 2023-06-22 Thomas Brihaye , Florent Delgrange , Youssouf Oualhadj , Mickael Randour

In studying randomized search heuristics, a frequent quantity of interest is the first time a (real-valued) stochastic process obtains (or passes) a certain value. The processes under investigation commonly show a bias towards this goal,…

Probability · Mathematics 2024-06-24 Timo Kötzing