English
Related papers

Related papers: A note on weak compactness of occupation measures …

200 papers

In proving large deviation estimates, the lower bound for open sets and upper bound for compact sets are essentially local estimates. On the other hand, the upper bound for closed sets is global and compactness of space or an exponential…

Probability · Mathematics 2015-10-20 Chiranjib Mukherjee , S. R. S. Varadhan

We propose solution methods for previously-unsolved constrained MDPs in which actions can continuously modify the transition probabilities within some acceptable sets. While many methods have been proposed to solve regular MDPs with large…

Artificial Intelligence · Computer Science 2013-09-27 Marek Petrik , Dharmashankar Subramanian , Janusz Marecki

This paper addresses the challenge of a particular class of noisy state observations in Markov Decision Processes (MDPs), a common issue in various real-world applications. We focus on modeling this uncertainty through a confusion matrix…

Machine Learning · Computer Science 2023-12-15 Amirhossein Afsharrad , Sanjay Lall

We consider positively supported Borel measures for which all moments exist. On the set of compactly supported measures in this class a partial order is defined via eventual dominance of the moment sequences. Special classes are identified…

Classical Analysis and ODEs · Mathematics 2022-03-23 Vincent Bürgin , Jeremias Epperlein , Fabian Wirth

This paper provides sufficient conditions over the sequence of samples and parameters of an adaptive Markov Chain Monte Carlo (MCMC) algorithm to ensure ergodicity with respect to a target distribution that can have unbounded support. These…

Statistics Theory · Mathematics 2026-02-17 Alexandre Chotard

We construct and study branching Markov processes on the space of finite configurations of the state space of a given standard process, controlled by a branching kernel and a killing one. In particular, we may start with a superprocess,…

Probability · Mathematics 2015-08-03 Lucian Beznea , Oana Lupascu

This paper is devoted to studying constrained continuous-time Markov decision processes (MDPs) in the class of randomized policies depending on state histories. The transition rates may be unbounded, the reward and costs are admitted to be…

Probability · Mathematics 2012-01-04 Xianping Guo , Xinyuan Song

Robust Markov Decision Processes (MDPs) are receiving much attention in learning a robust policy which is less sensitive to environment changes. There are an increasing number of works analyzing sample-efficiency of robust MDPs. However,…

Machine Learning · Statistics 2023-09-13 Wenhao Yang , Han Wang , Tadashi Kozuno , Scott M. Jordan , Zhihua Zhang

We examine a two-dimensional system of sterically repulsive interacting disks where each particle runs in a random direction. This system is equivalent to a run-and-tumble dynamics system in the limit where the run time is infinite. At low…

Soft Condensed Matter · Physics 2017-12-06 C. Reichhardt , C. J. Olson Reichhardt

In many engineering systems, proper predictive maintenance and operational control are essential to increase efficiency and reliability while reducing maintenance costs. However, one of the major challenges is that many sensors are used for…

Applications · Statistics 2025-12-09 Boyang Xu , Yunyi Kang , Xinyu Zhao , Hao Yan , Feng Ju

This paper studies Markov Decision Processes under parameter uncertainty. We adapt the distributionally robust optimization framework, and assume that the uncertain parameters are random variables following an unknown distribution, and…

Systems and Control · Computer Science 2015-05-14 Pengqian Yu , Huan Xu

We consider a collection of fully coupled weakly interacting diffusion processes moving in a two-scale environment. We study the moderate deviations principle of the empirical distribution of the particles' positions in the combined limit…

Probability · Mathematics 2023-07-17 Zachary Bezemek , Konstantinos Spiliopoulos

Low-Rank Markov Decision Processes (MDPs) have recently emerged as a promising framework within the domain of reinforcement learning (RL), as they allow for provably approximately correct (PAC) learning guarantees while also incorporating…

Machine Learning · Computer Science 2024-04-03 Andrew Bennett , Nathan Kallus , Miruna Oprescu

The focus of this article is on entropy and Markov processes. We study the properties of functionals which are invariant with respect to monotonic transformations and analyze two invariant "additivity" properties: (i) existence of a…

Data Analysis, Statistics and Probability · Physics 2013-11-12 A. N. Gorban , P. A. Gorban , G. Judge

Advances in mobile computing technologies have made it possible to monitor and apply data-driven interventions across complex systems in real time. Markov decision processes (MDPs) are the primary model for sequential decision problems with…

Methodology · Statistics 2018-03-20 Longshaokan Wang , Eric B. Laber , Katie Witkiewitz

We study the evaluation of a policy under best- and worst-case perturbations to a Markov decision process (MDP), using transition observations from the original MDP, whether they are generated under the same or a different policy. This is…

Artificial Intelligence · Computer Science 2024-11-05 Andrew Bennett , Nathan Kallus , Miruna Oprescu , Wen Sun , Kaiwen Wang

We consider a distributionally robust Partially Observable Markov Decision Process (DR-POMDP), where the distribution of the transition-observation probabilities is unknown at the beginning of each decision period, but their realizations…

Optimization and Control · Mathematics 2020-12-09 Hideaki Nakao , Ruiwei Jiang , Siqian Shen

This paper addresses a key limitation in existing counterfactual inference methods for Markov Decision Processes (MDPs). Current approaches assume a specific causal model to make counterfactuals identifiable. However, there are usually many…

Artificial Intelligence · Computer Science 2026-05-25 Jessica Lally , Milad Kazemi , Nicola Paoletti

Consider the lattice of bounded linear operators on the space of Borel measures on a Polish space. We prove that the operators which are continuous with respect to the weak topology induced by the bounded measurable functions form a…

Functional Analysis · Mathematics 2015-11-05 Moritz Gerlach , Markus Kunze

Stochastic and soft optimal policies resulting from entropy-regularized Markov decision processes (ER-MDP) are desirable for exploration and imitation learning applications. Motivated by the fact that such policies are sensitive with…

Machine Learning · Computer Science 2022-01-03 Tien Mai , Patrick Jaillet
‹ Prev 1 4 5 6 7 8 10 Next ›