English
Related papers

Related papers: Finite Horizon Decision Timing with Partially Obse…

200 papers

Autonomous agents often operate in scenarios where the state is partially observed. In addition to maximizing their cumulative reward, agents must execute complex tasks with rich temporal and logical structures. These tasks can be expressed…

Systems and Control · Electrical Eng. & Systems 2022-03-18 Krishna C. Kalagarla , Dhruva Kartik , Dongming Shen , Rahul Jain , Ashutosh Nayyar , Pierluigi Nuzzo

Suppose an online platform wants to compare a treatment and control policy, e.g., two different matching algorithms in a ridesharing system, or two different inventory management algorithms in an online retail site. Standard randomized…

Methodology · Statistics 2022-12-27 Peter Glynn , Ramesh Johari , Mohammad Rasouli

In this paper, based on real-time nonlinear receding horizon control methodology, a novel approach is developed for parameter estimation of time invariant and time varying nonlinear dynamical systems in chaotic environments. Here, the…

Optimization and Control · Mathematics 2016-11-21 Fei Sun , Kamran Turkoglu

How do decisions change with the economic environment and with time? This paper studies general nonstationary stopping problems and provides the methodological tools to answer these questions. First, we identify conditions that ensure a…

Theoretical Economics · Economics 2024-08-01 Théo Durandard , Matteo Camboni

This paper studies the problem of steering a linear time-invariant system subject to state and input constraints towards a goal location that may be inferred only through partial observations. We assume mixed-observable settings, where the…

Optimization and Control · Mathematics 2022-11-22 Ugo Rosolia , Yuxiao Chen , Shreyansh Daftry , Masahiro Ono , Yisong Yue , Aaron D. Ames

Bayesian optimization is a coherent, ubiquitous approach to decision-making under uncertainty, with applications including multi-arm bandits, active learning, and black-box optimization. Bayesian optimization selects decisions (i.e.…

Machine Learning · Computer Science 2023-12-13 Samuel Stanton , Wesley Maddox , Andrew Gordon Wilson

We consider impulse control problems in finite horizon for diffusions with decision lag and execution delay. The new feature is that our general framework deals with the important case when several consecutive orders may be decided before…

Probability · Mathematics 2007-05-23 Benjamin Bruder , Huyen Pham

Contextual bandits constitute a classical framework for decision-making under uncertainty. In this setting, the goal is to learn the arms of highest reward subject to contextual information, while the unknown reward parameters of each arm…

Machine Learning · Statistics 2024-02-19 Hongju Park , Mohamad Kazem Shirani Faradonbeh

In this paper, the trajectory planning problem for autonomous rendezvous and docking between a controlled spacecraft and a tumbling target is addressed. The use of a variable planning horizon is proposed in order to construct an appropriate…

Systems and Control · Electrical Eng. & Systems 2024-09-11 Mirko Leomanni , Renato Quartullo , Gianni Bianchini , Andrea Garulli , Antonio Giannitrapani

In this paper, we study a remote monitoring system where a receiver observes a remote binary Markov source and decides whether to sample and transmit the state through a randomly delayed channel. We adopt uncertainty of information (UoI),…

Information Theory · Computer Science 2024-05-20 Xiaomeng Chen , Aimin Li , Shaohua Wu

We study planning problems faced by robots operating in uncertain environments with incomplete knowledge of state, and actions that are noisy and/or imprecise. This paper identifies a new problem sub-class that models settings in which…

Robotics · Computer Science 2022-08-09 Federico Rossi , Dylan Shell

This paper investigates manipulation of multiple unknown objects in a crowded environment. Because of incomplete knowledge due to unknown objects and occlusions in visual observations, object observations are imperfect and action success is…

Robotics · Computer Science 2014-07-09 Joni Pajarinen , Ville Kyrki

The model consists of a signal process $X$ which is a general Brownian diffusion process and an observation process $Y$, also a diffusion process, which is supposed to be correlated to the signal process. We suppose that the process $Y$ is…

Probability · Mathematics 2012-11-20 Christophe Pofeta , Abass Sagna

This paper is concerned with a time-inconsistent stochastic optimal control problem in an infinite time horizon with a non-degenerate diffusion in the state equation. A major assumption is that people become rational after a large time.…

Optimization and Control · Mathematics 2025-09-19 Qingmeng Wei , Jiongmin Yong

Reward optimization in fully observable Markov decision processes is equivalent to a linear program over the polytope of state-action frequencies. Taking a similar perspective in the case of partially observable Markov decision processes…

Machine Learning · Computer Science 2022-05-30 Johannes Müller , Guido Montúfar

We consider a class of sequential decision-making problems under uncertainty that can encompass various types of supervised learning concepts. These problems have a completely observed state process and a partially observed modulation…

Optimization and Control · Mathematics 2021-08-24 R. Reid Bishop , Chelsea C. White

A resource-constrained system monitors a source of information by requesting a finite number of updates subject to random transmission delays. An a priori fixed update request policy is shown to minimize a polynomial penalty function of the…

Information Theory · Computer Science 2019-10-03 David Ramirez , Elza Erkip , H. Vincent Poor

This paper presents an optimization-based receding horizon trajectory planning algorithm for dynamical systems operating in unstructured and cluttered environments. The proposed approach is a two-step procedure that uses a motion planning…

Optimization and Control · Mathematics 2019-12-12 Kristoffer Bergman , Oskar Ljungqvist , Torkel Glad , Daniel Axehill

Autonomous systems often have logical constraints arising, for example, from safety, operational, or regulatory requirements. Such constraints can be expressed using temporal logic specifications. The system state is often partially…

Artificial Intelligence · Computer Science 2024-06-21 Krishna C. Kalagarla , Dhruva Kartik , Dongming Shen , Rahul Jain , Ashutosh Nayyar , Pierluigi Nuzzo

The aim of the paper is to extend the model of "fishing problem". The simple formulation is following. The angler goes to fishing. He buys fishing ticket for a fixed time. There are two places for fishing at the lake. The fishes are caught…

Probability · Mathematics 2008-12-19 Anna Karpowicz , Krzysztof Szajowski