English
Related papers

Related papers: Grab It Before It's Gone: Testing Uncertain Reward…

200 papers

We study finite horizon reachable set estimation for unknown discrete-time dynamical systems using only sampled state trajectories. Rather than treating scenario optimization as a black-box tool, we show how it can be tailored to reachable…

Optimization and Control · Mathematics 2026-04-15 Georgios Pantazis , Michelle S. Chong

One problem of wide interest involves estimating expected crossing-times. Several tools have been developed to solve this problem beginning with the works of Wald and the theory of sequential analysis. An extension of his approach is…

Methodology · Statistics 2015-06-17 Mark Brown , Victor de la Pena , Tony Sit

We propose a fast algorithm for the probabilistic solution of boundary value problems (BVPs), which are ordinary differential equations subject to boundary conditions. In contrast to previous work, we introduce a Gauss--Markov prior and…

Machine Learning · Statistics 2021-06-16 Nicholas Krämer , Philipp Hennig

We consider the problem of estimating the expected value of information (the knowledge gradient) for Bayesian learning problems where the belief model is nonlinear in the parameters. Our goal is to maximize some metric, while simultaneously…

Machine Learning · Statistics 2016-11-23 Xinyu He , Warren B. Powell

In this work, we address the output--feedback control problem for nonlinear systems under bounded disturbances using a moving horizon approach. The controller is posed as an optimization-based problem that simultaneously estimates the state…

Systems and Control · Electrical Eng. & Systems 2024-09-23 Nestor N. Deniz , Guido Sanchez , Marina H. Murillo , Leonardo L. Giovanini

The paper addresses the problem of computing maximal conditional expected accumulated rewards until reaching a target state (briefly called maximal conditional expectations) in finite-state Markov decision processes where the condition is…

Logic in Computer Science · Computer Science 2023-03-07 Christel Baier , Joachim Klein , Sascha Klüppelholz , Sascha Wunderlich

For the problem of task-agnostic reinforcement learning (RL), an agent first collects samples from an unknown environment without the supervision of reward signals, then is revealed with a reward and is asked to compute a corresponding…

Machine Learning · Computer Science 2022-03-16 Jingfeng Wu , Vladimir Braverman , Lin F. Yang

The socioeconomic impact of pollution naturally comes with uncertainty due to, e.g., current new technological developments in emissions' abatement or demographic changes. On top of that, the trend of the future costs of the environmental…

Optimization and Control · Mathematics 2024-02-28 Matteo Basei , Giorgio Ferrari , Neofytos Rodosthenous

We study the regularity of the stochastic representation of the solution of a class of initial-boundary value problems related to a regime-switching diffusion. This representation is related to the value function of a finite-horizon optimal…

Probability · Mathematics 2017-06-12 S. D. Jacka , A. Ocejo

This article is devoted to the stochastic anticipating equations with the extended stochastic integral with respect to the Gaussian processes of a special type. In the particular cases the solutions of such an equations are the well-known…

Probability · Mathematics 2007-05-23 Andrey A Dorogovtsev

We introduce a class of learning problems where the agent is presented with a series of tasks. Intuitively, if there is relation among those tasks, then the information gained during execution of one task has value for the execution of…

Machine Learning · Computer Science 2012-09-06 Christos Dimitrakakis

We study a class of infinite horizon impulse control problems with execution delay when the dynamics of the system is described by a general adapted stochastic process. The problem is solved by means of probabilistic tools relying on the…

Probability · Mathematics 2019-05-21 Boualem Djehiche , Said Hamadene , Ibtissem Hdhiri , Helmi Zaatra

We deal with an infinite horizon, infinite dimensional stochastic optimal control problem arising in the study of economic growth in time-space. Such problem has been the object of various papers in deterministic cases when the possible…

Optimization and Control · Mathematics 2022-03-14 Fausto Gozzi , Marta Leocata

In this study, we develop a stochastic optimal control approach with reinforcement learning structure to learn the unknown parameters appeared in the drift and diffusion terms of the stochastic differential equation. By choosing an…

Optimization and Control · Mathematics 2023-08-22 Shuzhen Yang

This paper studies identification and inference of the welfare gain that results from switching from one policy (such as the status quo policy) to another policy. The welfare gain is not point identified in general when data are obtained…

Econometrics · Economics 2022-07-12 Undral Byambadalai

Estimation of individual treatment effects is commonly used as the basis for contextual decision making in fields such as healthcare, education, and economics. However, it is often sufficient for the decision maker to have estimates of…

Machine Learning · Computer Science 2020-08-13 Maggie Makar , Fredrik D. Johansson , John Guttag , David Sontag

Neural networks make accurate predictions but often fail to provide reliable uncertainty estimates, especially under covariate distribution shifts between training and testing. To address this problem, we propose a Bayesian framework for…

Machine Learning · Statistics 2025-12-22 Yuli Slavutsky , David M. Blei

We model stochastic choice as environment-dependent switching among a small library of deterministic decision rules. A Random Rule Model generates menu-level choice probabilities via named, interpretable rules weighted by observable menu…

General Economics · Economics 2026-04-15 Avner Seror

This paper considers stochastic-constrained stochastic optimization where the stochastic constraint is to satisfy that the expectation of a random function is below a certain threshold. In particular, we study the setting where data samples…

Optimization and Control · Mathematics 2026-01-27 Yeongjong Kim , Dabeen Lee

We study bandit best-arm identification with arbitrary and potentially adversarial rewards. A simple random uniform learner obtains the optimal rate of error in the adversarial scenario. However, this type of strategy is suboptimal when the…

Machine Learning · Statistics 2026-04-17 Yasin Abbasi-Yadkori , Peter L. Bartlett , Victor Gabillon , Alan Malek , Michal Valko
‹ Prev 1 3 4 5 6 7 10 Next ›