English
Related papers

Related papers: Linear response and moderate deviations: hierarchi…

200 papers

We consider large random trees under Gibbs distributions and prove a Large Deviation Principle (LDP) for the distribution of degrees of vertices of the tree. The LDP rate function is given explicitly. An immediate consequence is a Law of…

Probability · Mathematics 2009-11-13 Yuri Bakhtin , Christine Heitsch

We introduce a framework for approximate analysis of Markov decision processes (MDP) with bounded-, unbounded-, and infinite-horizon properties. The main idea is to identify a "core" of an MDP, i.e., a subsystem where we provably remain…

Systems and Control · Electrical Eng. & Systems 2023-06-22 Jan Křetínský , Tobias Meggendorfer

Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study that used a Hidden Markov Model (HMM) to infer latent error regimes from surrogate-physics…

Machine Learning · Computer Science 2026-04-27 Annice Najafi , Shokoufeh Mirzaei

Sample average approximation--based stochastic dynamic programming (SDP) and model predictive control (MPC) are two different methods for approaching multistage stochastic optimization. In this paper we investigate the conditions under…

Optimization and Control · Mathematics 2026-02-10 Dominic S. T. Keehan , Andrew B. Philpott , Edward J. Anderson

The Pseudo-Marginal (PM) algorithm is a popular Markov chain Monte Carlo (MCMC) method used to sample from a target distribution when its density is inaccessible, but can be estimated with a non-negative unbiased estimator. Its performance…

Computation · Statistics 2025-09-30 Sarra Abaoubida , Mylène Bédard , Florian Maire

A uniform key renewal theorem is deduced from the uniform Blackwell's renewal theorem. A uniform LDP (large deviations principle) for renewal-reward processes is obtained, and MDP (moderate deviations principle) is deduced under conditions…

Probability · Mathematics 2012-07-06 Boris Tsirelson

Importance sampling has become an important tool for the computation of tail-based risk measures. Since such quantities are often determined mainly by rare events standard Monte Carlo can be inefficient and importance sampling provides a…

Probability · Mathematics 2013-06-29 Pierre Nyquist

We study a system of interacting particles that randomly react to form new particles. The reaction flux is the rescaled number of reactions that take place in a time interval. We prove a dynamic large-deviation principle for the reaction…

Probability · Mathematics 2019-10-02 Robert Patterson , Michiel Renger

The Markov decision process (MDP) formulation used to model many real-world sequential decision making problems does not efficiently capture the setting where the set of available decisions (actions) at each time step is stochastic.…

Machine Learning · Computer Science 2020-01-22 Yash Chandak , Georgios Theocharous , Blossom Metevier , Philip S. Thomas

This paper studies Markov Decision Processes (MDPs) with atomless initial state distributions and atomless transition probabilities. Such MDPs are called atomless. The initial state distribution is considered to be fixed. We show that for…

Optimization and Control · Mathematics 2018-10-26 Eugene A. Feinberg , Aleksey B. Piunovskiy

Relative Divergence (RD) and Maximum Relative Divergence Principle (MRDP) for grading (order-comonotonic) functions (GF) on posets are used as an expression of Insufficient Reason Principle under the given prior information (IRP+). Classic…

Information Theory · Computer Science 2025-10-07 Alexander Dukhovny

In this paper we prove large and moderate deviations principles for the recursive kernel estimator of a probability density function and its partial derivatives. Unlike the density estimator, the derivatives estimators exhibit a quadratic…

Statistics Theory · Mathematics 2007-06-13 Abdelkader Mokkadem , Mariane Pelletier , Baba Thiam

In the study of natural and artificial complex systems, responses that are not completely determined by the considered decision variables are commonly modelled probabilistically, resulting in response distributions varying across decision…

Methodology · Statistics 2021-10-07 Athénaïs Gautier , David Ginsbourger , Guillaume Pirot

In this paper, we investigate the concentration properties of cumulative reward in Markov Decision Processes (MDPs), focusing on both asymptotic and non-asymptotic settings. We introduce a unified approach to characterize reward…

Machine Learning · Computer Science 2025-12-04 Borna Sayedana , Peter E. Caines , Aditya Mahajan

In this paper we derive a Large Deviation Principle (LDP) for inhomogeneous U/V-statistics of a general order. Using this, we derive a LDP for two types of statistics: random multilinear forms, and number of monochromatic copies of a…

Probability · Mathematics 2026-04-01 Sohom Bhattacharya , Nabarun Deb , Sumit Mukherjee

Partially observable Markov decision processes (POMDPs) provide an elegant mathematical framework for modeling complex decision and planning problems in stochastic domains in which states of the system are observable only indirectly, via a…

Artificial Intelligence · Computer Science 2011-06-02 M. Hauskrecht

In offline reinforcement learning (RL), the absence of active exploration calls for attention on the model robustness to tackle the sim-to-real gap, where the discrepancy between the simulated and deployed environments can significantly…

Machine Learning · Computer Science 2024-06-28 He Wang , Laixi Shi , Yuejie Chi

This paper studies Markov Decision Processes under parameter uncertainty. We adapt the distributionally robust optimization framework, and assume that the uncertain parameters are random variables following an unknown distribution, and…

Systems and Control · Computer Science 2015-05-14 Pengqian Yu , Huan Xu

We establish a moderate deviation principle (MDP) for the number of eigenvalues of a Wigner matrix in an interval. The proof relies on fine asymptotics of the variance of the eigenvalue counting function of GUE matrices due to Gustavsson.…

Probability · Mathematics 2013-01-14 Hanna Doering , Peter Eichelsbacher

We consider a class of sequential decision-making problems under uncertainty that can encompass various types of supervised learning concepts. These problems have a completely observed state process and a partially observed modulation…

Optimization and Control · Mathematics 2021-08-24 R. Reid Bishop , Chelsea C. White