English
Related papers

Related papers: Finite Horizon Decision Timing with Partially Obse…

200 papers

In this paper, we study the remote estimation problem of a Markov process over a channel with a cost. We formulate this problem as an infinite horizon optimization problem with two players, i.e., a sensor and a monitor, that have distinct…

Systems and Control · Electrical Eng. & Systems 2024-02-01 Edoardo David Santi , Touraj Soleymani , Deniz Gunduz

We propose a moving horizon estimation scheme for estimating the states and time-varying parameters of nonlinear systems. We consider the case where observability of the parameters depends on the excitation of the system and may be absent…

Systems and Control · Electrical Eng. & Systems 2025-08-21 Julian D. Schiller , Matthias A. Müller

The partial monitoring (PM) framework provides a theoretical formulation of sequential learning problems with incomplete feedback. On each round, a learning agent plays an action while the environment simultaneously chooses an outcome. The…

Machine Learning · Computer Science 2024-05-17 Maxime Heuillet , Ola Ahmad , Audrey Durand

We revisit closed-loop performance guarantees for Model Predictive Control in the deterministic and stochastic cases, which extend to novel performance results applicable to receding horizon control of Partially Observable Markov Decision…

Optimization and Control · Mathematics 2020-05-01 Martin A. Sehr , Robert R. Bitmead

Understanding how people allocate visual attention is central to Human-Computer Interaction (HCI), yet existing computational models of attention are often either descriptive, task-specific, or difficult to interpret. My dissertation…

Human-Computer Interaction · Computer Science 2026-03-03 Yunpeng Bai

An iterative learning algorithm is presented for continuous-time linear-quadratic optimal control problems where the system is externally symmetric with unknown dynamics. Both finite-horizon and infinite-horizon problems are considered. It…

Optimization and Control · Mathematics 2025-10-10 Hamed Taghavian , Florian Dorfler , Mikael Johansson

We provide a framework for speeding up algorithms for time-bounded reachability analysis of continuous-time Markov decision processes. The principle is to find a small, but almost equivalent subsystem of the original system and only analyse…

Systems and Control · Computer Science 2018-07-26 Pranav Ashok , Yuliya Butkova , Holger Hermanns , Jan Křetínský

Bayesian optimization is a methodology to optimize black-box functions. Traditionally, it focuses on the setting where you can arbitrarily query the search space. However, many real-life problems do not offer this flexibility; in…

Zero-sum Dynkin games under Poisson constraints, where players can only stop at the event times of a Poisson process, have been studied widely in the recent literature. The constraint can be modelled in two ways: either both players share…

Optimization and Control · Mathematics 2025-12-09 David Hobson , Gechun Liang , Edward Wang

We study model-based learning of finite-window policies in tabular partially observable Markov decision processes (POMDPs). A common approach to learning under partial observability is to approximate unbounded history dependencies using…

Machine Learning · Computer Science 2026-04-02 Philip Jordan , Maryam Kamgarpour

In this paper we study the problem of information sharing among rational self-interested agents as a dynamic game of asymmetric information. We assume that the agents imperfectly observe a Markov chain and they are called to decide whether…

Computer Science and Game Theory · Computer Science 2021-03-30 Konstantinos Ntemos , George Pikramenos , Nicholas Kalouptsidis

We investigate Markovian queues that are examined by a controller at random times determined by a Poisson process. Upon examination, the controller sets the service speed to be equal to the minimum of the current number of customers in the…

Performance · Computer Science 2023-03-30 R. Núñez-Queija , B. J. Prabhu , J. A. C. Resing

In this paper, we study the well-known team orienteering problem where a fleet of robots collects rewards by visiting locations. Usually, the rewards are assumed to be known to the robots; however, in applications such as environmental…

Robotics · Computer Science 2021-12-16 Nils Wilde , Armin Sadeghi , Stephen L. Smith

We propose a novel Bayesian method to solve the maximization of a time-dependent expensive-to-evaluate stochastic oracle. We are interested in the decision that maximizes the oracle at a finite time horizon, given a limited budget of noisy…

Computation · Statistics 2021-05-21 S. Ashwin Renganathan , Jeffrey Larson , Stefan M. Wild

We consider a discounted infinite horizon optimal stopping problem. If the underlying distribution is known a priori, the solution of this problem is obtained via dynamic programming (DP) and is given by a well known threshold rule. When…

Machine Learning · Computer Science 2021-02-23 Daniel Russo , Assaf Zeevi , Tianyi Zhang

This paper considers the optimal dividend payment problem in piecewise-deterministic compound Poisson risk models. The objective is to maximize the expected discounted dividend payout up to the time of ruin. We provide a comparative study…

Optimization and Control · Mathematics 2016-08-02 Runhuan Feng , Hans Volkmer , Shuaiqi Zhang , Chao Zhu

We present a numerical method to compute expectations of functionals of a piecewise-deterministic Markov process. We discuss time dependent functionals as well as deterministic time horizon problems. Our approach is based on the…

Probability · Mathematics 2012-01-31 Adrien Brandejsky , Benoîte de Saporta , François Dufour

Autonomous systems are often required to operate in partially observable environments. They must reliably execute a specified objective even with incomplete information about the state of the environment. We propose a methodology to…

Artificial Intelligence · Computer Science 2020-01-14 Maxime Bouton , Jana Tumova , Mykel J. Kochenderfer

We present the theoretical analysis and proofs of a recently developed algorithm that allows for optimal planning over long and infinite horizons for achieving multiple independent tasks that are partially observable and evolve over time.

Robotics · Computer Science 2021-02-26 Anahita Mohseni-Kabir , Manuela Veloso , Maxim Likhachev

We consider a bivariate first hitting-time model in which durations are the crossing times of dependent compound Poisson processes with fixed thresholds. The identifiability of the model is discussed, and likelihood estimators of the model…

Methodology · Statistics 2025-04-14 Mikael Escobar-Bach , Alexandre Popier , Malo Sahin
‹ Prev 1 4 5 6 7 8 10 Next ›