English
Related papers

Related papers: Linear response and moderate deviations: hierarchi…

200 papers

We investigate the parameter recovery of Markov-switching ordinary differential processes from discrete observations, where the differential equations are nonlinear additive models. This framework has been widely applied in biological…

Methodology · Statistics 2025-01-03 Katherine Tsai , Mladen Kolar , Sanmi Koyejo

This paper considers an infinite-horizon Markov decision process (MDP) that allows for general non-exponential discount functions, in both discrete and continuous time. Due to the inherent time inconsistency, we look for a randomized…

Optimization and Control · Mathematics 2024-12-10 Erhan Bayraktar , Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

We consider Markov decision processes (MDPs) in which the transition probabilities and rewards belong to an uncertainty set parametrized by a collection of random variables. The probability distributions for these random parameters are…

Logic in Computer Science · Computer Science 2020-02-26 Murat Cubuktepe , Nils Jansen , Sebastian Junges , Joost-Pieter Katoen , Ufuk Topcu

A class of discrete distributions can be derived from stationary renewal processes. They have the useful property that the mean is a simple function of the model parameters. Thus regressions of the distribution mean on covariates can be…

Methodology · Statistics 2018-03-01 Rose Baker

In this paper, we present a novel framework to synthesize robust strategies for discrete-time nonlinear systems with random disturbances that are unknown, against temporal logic specifications. The proposed framework is data-driven and…

Systems and Control · Electrical Eng. & Systems 2025-04-29 Ibon Gracia , Luca Laurenti , Manuel Mazo , Alessandro Abate , Morteza Lahijanian

The aim of this paper is to get asymptotic deviation bounds via a Large Deviation Principle (LDP) for cumulative processes also known as compound renewal processes or renewal-reward processes. These processes cumulate independent random…

Probability · Mathematics 2023-06-21 Patrick Cattiaux , Laetitia Colombani , Manon Costa

We cast episodic Markov decision process (MDP) planning as Bayesian inference over policies. A policy is treated as the latent variable and is assigned an unnormalized probability of optimality that is monotone in its expected return,…

Machine Learning · Computer Science 2026-04-14 David Tolpin

Value Iteration is a widely used algorithm for solving Markov Decision Processes (MDPs). While previous studies have extensively analyzed its convergence properties, they primarily focus on convergence with respect to the infinity norm. In…

Machine Learning · Computer Science 2025-02-06 Arsenii Mustafin , Sebastien Colla , Alex Olshevsky , Ioannis Ch. Paschalidis

We consider a $\mathbb{R}^d$-valued branching random walk with a stationary and ergodic environment $\xi=(\xi_n)$ indexed by time $n\in\mathbb{N}$. Let $Z_n$ be the counting measure of particles of generation $n$. With the help of the…

Probability · Mathematics 2019-10-15 Chunmao Huang , Xin Wang , Xiaoqiang Wang

In this paper we prove large and moderate deviations principles for the recursive kernel estimator of a probability density function and its partial derivatives. Unlike the density estimator, the derivatives estimators exhibit a quadratic…

Statistics Theory · Mathematics 2007-06-13 Abdelkader Mokkadem , Mariane Pelletier , Baba Thiam

We consider Markov decision processes (MDPs) which are a standard model for probabilistic systems. We focus on qualitative properties for MDPs that can express that desired behaviors of the system arise almost-surely (with probability 1) or…

Logic in Computer Science · Computer Science 2014-05-06 Krishnendu Chatterjee , Martin Chmelik , Przemyslaw Daca

This paper studies Markov Decision Processes under parameter uncertainty. We adapt the distributionally robust optimization framework, and assume that the uncertain parameters are random variables following an unknown distribution, and…

Systems and Control · Computer Science 2015-05-14 Pengqian Yu , Huan Xu

We consider a distributionally robust Partially Observable Markov Decision Process (DR-POMDP), where the distribution of the transition-observation probabilities is unknown at the beginning of each decision period, but their realizations…

Optimization and Control · Mathematics 2020-12-09 Hideaki Nakao , Ruiwei Jiang , Siqian Shen

Cram\'er type moderate deviation theorems quantify the accuracy of the relative error of the normal approximation and provide theoretical justifications for many commonly used methods in statistics. In this paper, we develop a new…

Probability · Mathematics 2016-06-07 Qi-Man Shao , Wen-Xin Zhou

Robust Markov decision processes (MDPs) have attracted significant interest due to their ability to protect MDPs from poor out-of-sample performance in the presence of ambiguity. In contrast to classical MDPs, which account for…

Optimization and Control · Mathematics 2026-02-06 Chin Pang Ho , Marek Petrik , Wolfram Wiesemann

We consider the small deviation probabilities (SDP) for sums of stationary Gaussian sequences. For the cases of constant boundaries and boundaries tending to zero, we obtain quite general results. For the case of the boundaries tending to…

Probability · Mathematics 2020-02-11 Frank Aurzada , Mikhail Lifshits

Markov decision processes (MDPs) are a fundamental model for decision making under uncertainty. They exhibit non-deterministic choice as well as probabilistic uncertainty. Traditionally, verification algorithms assume exact knowledge of the…

Artificial Intelligence · Computer Science 2025-04-18 Tobias Meggendorfer , Maximilian Weininger , Patrick Wienhöft

We consider risk-sensitive Markov decision processes (MDPs), where the MDP model is influenced by a parameter which takes values in a compact metric space. We identify sufficient conditions under which small perturbations in the model…

Optimization and Control · Mathematics 2022-09-28 Shiping Shao , Abhishek Gupta , William B. Haskell

We prove moderate deviation principles for the tagged particle position and current in one-dimensional symmetric simple exclusion processes. There is at most one particle per site. A particle jumps to one of its two neighbors at rate $1/2$,…

Probability · Mathematics 2022-03-11 Xiaofeng Xue , Linjie Zhao

Differential Dynamic Programming (DDP) is an efficient computational tool for solving nonlinear optimal control problems. It was originally designed as a single shooting method and thus is sensitive to the initial guess supplied. This work…

Robotics · Computer Science 2023-09-29 He Li , Wenhao Yu , Tingnan Zhang , Patrick M. Wensing