English
Related papers

Related papers: Introducing user-prescribed constraints in Markov …

200 papers

We develop algorithms with low regret for learning episodic Markov decision processes based on kernel approximation techniques. The algorithms are based on both the Upper Confidence Bound (UCB) as well as Posterior or Thompson Sampling…

Machine Learning · Computer Science 2019-11-06 Sayak Ray Chowdhury , Aditya Gopalan

We present a case study applying learning-based distributionally robust model predictive control to highway motion planning under stochastic uncertainty of the lane change behavior of surrounding road users. The dynamics of road users are…

Systems and Control · Electrical Eng. & Systems 2022-11-08 Mathijs Schuurmans , Alexander Katriniok , Christopher Meissen , H. Eric Tseng , Panagiotis Patrinos

Stochastic simulation models are generative models that mimic complex systems to help with decision-making. The reliability of these models heavily depends on well-calibrated input model parameters. However, in many practical scenarios,…

Methodology · Statistics 2024-11-11 Ziwei Su , Diego Klabjan

For a large class of Markov Decision Processes, stationary (possibly randomized) policies are globally optimal. However, in Borel state and action spaces, the computation and implementation of even such stationary policies are known to be…

Optimization and Control · Mathematics 2014-04-29 Naci Saldi , Tamás Linder , Serdar Yüksel

We propose an algorithm for designing optimal inputs for on-line Bayesian identification of stochastic non-linear state-space models. The proposed method relies on minimization of the posterior Cram\'er Rao lower bound derived for the model…

Applications · Statistics 2013-07-25 Aditya Tulsyan , Swanand R. Khare , Biao Huang , R. Bhushan Gopaluni , J. Fraser Forbes

The development of algorithms for unsupervised pattern recognition by nonlinear clustering is a notable problem in data science. Markov clustering (MCL) is a renowned algorithm that simulates stochastic flows on a network of sample…

Machine Learning · Computer Science 2019-12-30 C. Duran , A. Acevedo , S. Ciucci , A. Muscoloni , CV. Cannistraci

We present several Monte Carlo strategies for simulating discrete-time Markov chains with continuous multi-dimensional state space; we focus on stratified techniques. We first analyze the variance of the calculation of the measure of a…

Statistics Theory · Mathematics 2016-03-22 Rana Fakhereddine , Rami El Haddad , Christian Lécot

This papers deals with the constrained discounted control of piecewise deterministic Markov process (PDMPs) in general Borel spaces. The control variable acts on the jump rate and transition measure, and the goal is to minimize the total…

Optimization and Control · Mathematics 2014-02-26 Oswaldo Costa , François Dufour

We consider continuous-space, discrete-time Markov chains on $\mathbb{R}^d$, that admit a finite number $N$ of metastable states. Our main motivation for investigating these processes is to analyse random Poincar\'e maps, which describe…

Probability · Mathematics 2025-08-19 Nils Berglund

Optimal sensor scheduling with applications to networked estimation and control systems is considered. We model sensor measurement and transmission instances using jumps between states of a continuous-time Markov chain. We introduce a cost…

Optimization and Control · Mathematics 2014-05-07 Farhad Farokhi , Karl H. Johansson

This article describes a method for computing limits of a class of non-stationary Markov chains motivated by healthcare sojourn-time cycles. A mathematical validation of the computation method is also given. Applications are described that…

Probability · Mathematics 2024-11-19 Samuel Awoniyi

The preparation of the stationary distribution of irreducible, time-reversible Markov chains is a fundamental building block in many heuristic approaches to algorithmically hard problems. It has been conjectured that quantum analogs of…

Quantum Physics · Physics 2015-02-20 Vedran Dunjko , Hans J. Briegel

We present a novel probabilistic approach for optimal path experimental design. In this approach a discrete path optimization problem is defined on a static navigation mesh, and trajectories are modeled as random variables governed by a…

Optimization and Control · Mathematics 2026-01-19 Ahmed Attia

We consider the linear programming approach for constrained and unconstrained Markov decision processes (MDPs) under the long-run average cost criterion, where the class of MDPs in our study have Borel state spaces and discrete countable…

Optimization and Control · Mathematics 2021-04-20 Huizhen Yu

In most adaptive signal processing applications, system linearity is assumed and adaptive linear filters are thus used. The traditional class of supervised adaptive filters rely on error-correction learning for their adaptive capability.…

Machine Learning · Computer Science 2015-08-31 Songlin Zhao

Markov Chain Monte Carlo (MCMC) techniques are now widely used for cosmological parameter estimation. Chains are generated to sample the posterior probability distribution obtained following the Bayesian approach. An important issue is how…

In this paper, selection of an active sensor subset for tracking a discrete time, finite state Markov chain having an unknown transition probability matrix (TPM) is considered. A total of N sensors are available for making observations of…

Machine Learning · Computer Science 2020-11-02 Mrigank Raman , Ojal Kumar , Arpan Chattopadhyay

Behavioural metrics have been shown to be an effective mechanism for constructing representations in reinforcement learning. We present a novel perspective on behavioural metrics for Markov decision processes via the use of positive…

Machine Learning · Computer Science 2023-11-01 Pablo Samuel Castro , Tyler Kastner , Prakash Panangaden , Mark Rowland

We prove a central limit theorem for a general class of adaptive Markov Chain Monte Carlo algorithms driven by sub-geometrically ergodic Markov kernels. We discuss in detail the special case of stochastic approximation. We use the result to…

Probability · Mathematics 2009-11-03 Yves F. Atchade , Gersende Fort

We present an empirical, gradient-based method for solving data-driven stochastic optimal control problems using the theory of kernel embeddings of distributions. By embedding the integral operator of a stochastic kernel in a reproducing…

Optimization and Control · Mathematics 2022-09-20 Adam J. Thorpe , Jake A. Gonzales , Meeko M. K. Oishi