English
Related papers

Related papers: A note on weak compactness of occupation measures …

200 papers

We consider Markov Decision Processes (MDPs) in which every stationary policy induces the same graph structure for the underlying Markov chain and further, the graph has the following property: if we replace each recurrent class by a node,…

Machine Learning · Computer Science 2021-03-10 Joseph Lubars , Anna Winnicki , Michael Livesay , R. Srikant

Doubly robust methods hold considerable promise for off-policy evaluation in Markov decision processes (MDPs) under sequential ignorability: They have been shown to converge as $1/\sqrt{T}$ with the horizon $T$, to be statistically…

Machine Learning · Statistics 2025-09-30 Mohammad Mehrabi , Stefan Wager

This paper provides new sufficient conditions so that the optimal policy of a partially observed Markov decision process (POMDP) can be lower bounded by a myopic policy. The two new proposed conditions, namely, Lehmann precision and…

Systems and Control · Computer Science 2018-10-23 Vikram Krishnamurthy

Incorporating time-varying elements into electromagnetic systems has shown to be a powerful approach to challenge well-established performance limits, for example bounds on absorption and impedance matching. So far, the majority of these…

Optics · Physics 2024-04-09 Zeki Hayran , Francesco Monticone

This paper addresses the problem of optimal control of robotic sensing systems aimed at autonomous information gathering in scenarios such as environmental monitoring, search and rescue, and surveillance and reconnaissance. The information…

Systems and Control · Computer Science 2016-01-28 Mikko Lauri , Nikolay Atanasov , George J. Pappas , Risto Ritala

The notion of $\Delta$-weakly mixing set is introduced, which shares similar properties of weakly mixing sets. It is shown that if a dynamical system has positive topological entropy, then the collection of $\Delta$-weakly mixing sets is…

Dynamical Systems · Mathematics 2016-11-08 Wen Huang , Jian Li , Xiangdong Ye , Xiaoyao Zhou

We propose and analyze a temporal concatenation heuristic for solving large-scale finite-horizon Markov decision processes (MDP), which divides the MDP into smaller sub-problems along the time horizon and generates an overall solution by…

Optimization and Control · Mathematics 2022-06-22 Ruiyang Song , Kuang Xu

The aim of this paper is to prove ergodic decomposition theorems for probability measures quasi-invariant under Borel actions of inductively compact groups (Theorem 1) as well as for sigma-finite invariant measures (Corollary 1). For…

Dynamical Systems · Mathematics 2014-07-28 Alexander I. Bufetov

Simple random coverage models, well studied in Euclidean space, can also be defined on a general compact metric space. By analogy with the geometric models, and with the discrete coupon collector's problem and with cover times for finite…

Probability · Mathematics 2021-02-01 David J. Aldous

For time-inconsistent stochastic controls in discrete time and finite horizon, an open problem in Bj\"ork and Murgoci (Finance Stoch, 2014) is the existence of an equilibrium control. A nonrandomized Borel measurable Markov equilibrium…

Optimization and Control · Mathematics 2023-12-18 Erhan Bayraktar , Bingyan Han

We study continuity and discontinuity of the upper and lower (modified) box-counting, Hausdorff, packing, (modified) correlation measure-dimension mappings under the weak, setwise and TV topology on the space of Borel measures respectively…

Dynamical Systems · Mathematics 2021-05-13 Liangang Ma

In this note we study a natural measure on plane partitions giving rise to a certain discrete-time Muttalib-Borodin process (MBP): each time-slice is a discrete version of a Muttalib-Borodin ensemble (MBE). The process is determinantal with…

Probability · Mathematics 2020-10-30 Dan Betea , Alessandra Occelli

This paper provides conditions under which total-cost and average-cost Markov decision processes (MDPs) can be reduced to discounted ones. Results are given for transient total-cost MDPs with tran- sition rates whose values may be greater…

Optimization and Control · Mathematics 2017-05-04 Eugene A. Feinberg , Jefferson Huang

We consider synchronizing properties of Markov decision processes (MDP), viewed as generators of sequences of probability distributions over states. A probability distribution is p-synchronizing if the probability mass is at least p in some…

Logic in Computer Science · Computer Science 2014-07-01 Laurent Doyen , Thierry Massart , Mahsa Shirmohammadi

This paper studies the synthesis of a joint control and active perception policy for a stochastic system modeled as a partially observable Markov decision process (POMDP), subject to temporal logic specifications. The POMDP actions…

Systems and Control · Electrical Eng. & Systems 2025-04-21 Chongyang Shi , Michael R. Dorothy , Jie Fu

In this work, we study the problem of actively classifying the attributes of dynamical systems characterized as a finite set of Markov decision process (MDP) models. We are interested in finding strategies that actively interact with the…

Systems and Control · Electrical Eng. & Systems 2023-01-06 Bo Wu , Niklas Lauffer , Mohamadreza Ahmadi , Suda Bharadwaj , Zhe Xu , Ufuk Topcu

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

Optimization and Control · Mathematics 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang

We introduce Multi-Environment Markov Decision Processes (MEMDPs) which are MDPs with a set of probabilistic transition functions. The goal in a MEMDP is to synthesize a single controller with guaranteed performances against all…

Logic in Computer Science · Computer Science 2014-12-04 Jean-François Raskin , Ocan Sankur

Assume that $T$ is a conservative ergodic measure preserving transformation of the infinite measure space $(X,\mathcal{A},\mu)$.We study the asymptotic behaviour of occupation times of certain subsets of infinite measure. Specifically, we…

Dynamical Systems · Mathematics 2007-05-23 Jon Aaronson , Maximilian Thaler , Roland Zweimueller

Based on a weak convergence argument, we provide a necessary and sufficient condition that guarantees that a nonnegative local martingale is indeed a martingale. Typically, conditions of this sort are expressed in terms of integrability…

Probability · Mathematics 2014-04-24 Jose Blanchet , Johannes Ruf
‹ Prev 1 8 9 10 Next ›