English
Related papers

Related papers: Linear response and moderate deviations: hierarchi…

200 papers

The mean absolute deviation about the mean is an alternative to the standard deviation for measuring dispersion in a sample or in a population. For stationary, ergodic time series with a finite first moment, an asymptotic expansion for the…

Methodology · Statistics 2014-06-18 Johan Segers

Robust Markov Decision Processes (MDPs) are a powerful framework for modeling sequential decision-making problems with model uncertainty. This paper proposes the first first-order framework for solving robust MDPs. Our algorithm interleaves…

Optimization and Control · Mathematics 2021-01-18 Julien Grand-Clément , Christian Kroer

We introduce a new framework of episodic tabular Markov decision processes (MDPs) with adversarial preferences, which we refer to as preference-based MDPs (PbMDPs). Unlike standard episodic MDPs with adversarial losses, where the numerical…

Machine Learning · Computer Science 2025-07-17 Taira Tsuchiya , Shinji Ito , Haipeng Luo

For discretisations of hyperbolic conservation laws, mimicking properties of operators or solutions at the continuous (differential equation) level discretely has resulted in several successful methods. While well-posedness for nonlinear…

Numerical Analysis · Mathematics 2019-10-22 Hendrik Ranocha

Contextual MDPs are powerful tools with wide applicability in areas from biostatistics to machine learning. However, specializing them to offline datasets has been challenging due to a lack of robust, theoretically backed methods. Our work…

Machine Learning · Statistics 2026-05-06 Riddhiman Bhattacharyya , Sayak Chakrabarty , Imon Banerjee

Non-stationary environments are challenging for reinforcement learning algorithms. If the state transition and/or reward functions change based on latent factors, the agent is effectively tasked with optimizing a behavior that maximizes…

Machine Learning · Computer Science 2021-05-21 Lucas N. Alegre , Ana L. C. Bazzan , Bruno C. da Silva

The median probability model (MPM) Barbieri and Berger (2004) is defined as the model consisting of those variables whose marginal posterior probability of inclusion is at least 0.5. The MPM rule yields the best single model for prediction…

Statistics Theory · Mathematics 2018-08-20 Marilena Barbieri , James O. Berger , Edward I. George , Veronika Rockova

The problem of model selection is inevitable in an increasingly large number of applications involving partial theoretical knowledge and vast amounts of information, like in medicine, biology or economics. The associated techniques are…

Methodology · Statistics 2015-11-17 Stephane Guerrier , Maria-Pia Victoria-Feser

We investigate random walks in independent, identically distributed random sceneries under the assumption that the scenery variables satisfy Cramer's condition. We prove moderate deviation principles in dimensions two and larger, covering…

Probability · Mathematics 2007-05-23 Klaus Fleischmann , Peter Morters , Vitali Wachtel

When testing for the mean vector in a high dimensional setting, it is generally assumed that the observations are independently and identically distributed. However if the data are dependent, the existing test procedures fail to preserve…

Statistics Theory · Mathematics 2014-11-17 Deepak Nag Ayyala , Junyong Park , Anindya Roy

This paper addresses finite sample stability properties of sequential Monte Carlo methods for approximating sequences of probability distributions. The results presented herein are applicable in the scenario where the start and end…

Computation · Statistics 2015-03-19 Nick Whiteley

Maximum Mean Discrepancy (MMD) has been widely used in the areas of machine learning and statistics to quantify the distance between two distributions in the $p$-dimensional Euclidean space. The asymptotic property of the sample MMD has…

Statistics Theory · Mathematics 2023-08-29 Hanjia Gao , Xiaofeng Shao

We present a new geometric interpretation of Markov Decision Processes (MDPs) with a natural normalization procedure that allows us to adjust the value function at each state without altering the advantage of any action with respect to any…

Machine Learning · Computer Science 2025-03-06 Arsenii Mustafin , Aleksei Pakharev , Alex Olshevsky , Ioannis Ch. Paschalidis

In this work, we consider the regret minimization problem for reinforcement learning in latent Markov Decision Processes (LMDP). In an LMDP, an MDP is randomly drawn from a set of $M$ possible MDPs at the beginning of the interaction, but…

Machine Learning · Computer Science 2021-02-10 Jeongyeol Kwon , Yonathan Efroni , Constantine Caramanis , Shie Mannor

We establish a moderate deviations principle (MDP) for the log-determinant $\log | \det (M_n) |$ of a Wigner matrix $M_n$ matching four moments with either the GUE or GOE ensemble. Further we establish Cram\'er--type moderate deviations and…

Probability · Mathematics 2013-01-16 Hanna Döring , Peter Eichelsbacher

We provide a general approach to obtain upper bounds for small deviations $ \mathbb{P}(\Vert y \Vert \le \epsilon)$ in different norms, namely the supremum and $\beta$- H\"older norms. The large class of processes $y$ under consideration…

Probability · Mathematics 2015-02-18 Ehsan Azmoodeh , Lauri Viitasaari

We consider the problem of bounding large deviations for non-i.i.d. random variables that are allowed to have arbitrary dependencies. Previous works typically assumed a specific dependence structure, namely the existence of independent…

Probability · Mathematics 2018-11-06 Christoph H. Lampert , Liva Ralaivola , Alexander Zimin

POMDPs are useful models for systems where the true underlying state is not known completely to an outside observer; the outside observer incompletely knows the true state of the system, and observes a noisy version of the true system…

Machine Learning · Computer Science 2021-09-01 Caleb M. Bowyer

We show that for local alternatives to uniformity which are determined by a sequence of square integrable densities the moderate deviation (MD) theorem for the corresponding Neyman-Pearson statistic does not hold in the full range for all…

Statistics Theory · Mathematics 2020-03-27 Tadeusz Inglot

In this paper we show how to extend the Sample-Path Large Deviation Principle for the urn model of Hill, Lane and Sudderth to the case in which the increment of the urn is not a binary variable. In particular, we sketch how to modify the…

Probability · Mathematics 2025-11-19 Simone Franchini
‹ Prev 1 4 5 6 7 8 10 Next ›