English
Related papers

Related papers: Large Deviations in Renewal Models of Statistical …

200 papers

The analysis of structure-preserving numerical methods for the Poisson--Nernst--Planck (PNP) system has attracted growing interests in recent years. In this work, we provide an optimal rate convergence analysis and error estimate for finite…

Numerical Analysis · Mathematics 2022-02-23 Jie Ding , Cheng Wang , Shenggao Zhou

In a broad class of reinforcement learning applications, stochastic rewards have heavy-tailed distributions, which lead to infinite second-order moments for stochastic (semi)gradients in policy evaluation and direct policy optimization. In…

Machine Learning · Computer Science 2023-06-21 Semih Cayci , Atilla Eryilmaz

For renewal-reward processes with a power-law decaying waiting time distribution, anomalously large probabilities are assigned to atypical values of the asymptotic processes. Previous works have reveals that this anomalous scaling causes a…

Statistical Mechanics · Physics 2022-10-05 Hiroshi Horii , Raphael Lefevere , Masato Itami , Takahiro Nemoto

Reward modeling is central to alignment pipelines such as RLHF, RLAIF, and PPO-based policy optimization, yet its reliability is constrained by limited and heterogeneous human preference data that are expensive to collect at scale. While…

Machine Learning · Computer Science 2026-05-26 Payel Bhattacharjee , Osvaldo Simeone , Ravi Tandon

One dimensional pinning models have been widely studied in the physical and mathematical literature, also in presence of disorder. Roughly speaking, they undergo a transition between a delocalized phase and a localized one. In mathematical…

Mathematical Physics · Physics 2020-12-02 Giambattista Giacomin , Benjamin Havret

We study a class of dissipative PDE's perturbed by a bounded random kick force. It is assumed that the random force is non-degenerate, so that the Markov process obtained by the restriction of solutions to integer times has a unique…

Analysis of PDEs · Mathematics 2012-12-05 Vojkan Jaksic , Vahagn Nersesyan , Claude-Alain Pillet , Armen Shirikyan

Stochastic contraction analysis is a recently developed tool for studying the global stability properties of nonlinear stochastic systems, based on a differential analysis of convergence in an appropriate metric. To date, stochastic…

Optimization and Control · Mathematics 2013-04-02 Quang-Cuong Pham , Jean-Jacques Slotine

A large deviation principle is established for a two-scale stochastic system in which the slow component is a continuous process given by a small noise finite dimensional It\^{o} stochastic differential equation, and the fast component is a…

Probability · Mathematics 2017-05-09 Amarjit Budhiraja , Paul Dupuis , Arnab Ganguly

We prove pathwise large deviation principles of slow variables in slow-fast systems in the limit of time-scale separation tending to infinity. In the limit regime we consider, the convergence of the slow variable to its deterministic limit…

Probability · Mathematics 2020-11-25 Richard C. Kraaij , Mikola C. Schlottke

This study in centered on models accounting for stochastic deformations of sample paths of random walks, embedded either in $\mathbb{Z}^2$ or in $\mathbb{Z}^3$. These models are immersed in multi-type particle systems with exclusion.…

Statistical Mechanics · Physics 2007-05-23 Guy Fayolle , Cyril Furtlehner

A common and effective method for calculating the steady-state distribution of a process under stochastic resetting is the renewal approach that requires only the knowledge of the reset-free propagator of the underlying process and the…

Statistical Mechanics · Physics 2024-11-15 Ron Vatash , Amy Altshuler , Yael Roichman

We study discrete statistical mechanics systems perturbed by a random environment without a finite second moment. Specifically, we consider a random environment whose tail distribution satisfies $P[\omega > x] \sim x^{-\gamma}$ as $x \to…

Probability · Mathematics 2026-02-05 Gaspard Gomez

We propose to derive deviation measures through the Minkowski gauge of a given set of acceptable positions. We show that, given a suitable acceptance set, any positive homogeneous deviation measure can be accommodated in our framework. In…

Risk Management · Quantitative Finance 2021-07-27 Marlon Moresco , Marcelo Righi , Eduardo Horta

Reward modeling in large language models is susceptible to reward hacking, causing models to latch onto superficial features such as the tendency to generate lists or unnecessarily long responses. In reinforcement learning from human…

Computation and Language · Computer Science 2025-02-19 Taneesh Gupta , Shivam Shandilya , Xuchao Zhang , Rahul Madhavan , Supriyo Ghosh , Chetan Bansal , Huaxiu Yao , Saravan Rajmohan

We consider models of directed polymers interacting with a one-dimensional defect line on which random charges are placed. More abstractly, one starts from renewal sequence on $\Z$ and gives a random (site-dependent) reward or penalty to…

Probability · Mathematics 2007-06-13 F. L. Toninelli

We study classical stochastic systems with discrete states, coupled to switching external environments. For fast environmental processes we derive reduced dynamics for the system itself, focusing on corrections to the adiabatic limit of…

Statistical Mechanics · Physics 2019-03-27 Peter G. Hufton , Yen Ting Lin , Tobias Galla

We study delay-independent stability in nonlinear models with a distributed delay which have a positive equilibrium. Such models frequently occur in population dynamics and other applications. In particular, we construct a relevant…

Dynamical Systems · Mathematics 2009-01-12 Elena Braverman , Sergey Zhukovskiy

This paper introduces novel frameworks for large deviations and metastability analysis in heavy-tailed stochastic dynamical systems. We develop and apply these frameworks within the context of stochastic difference equation $X^\eta_{j+1}(x)…

Probability · Mathematics 2024-12-12 Xingyu Wang , Chang-Han Rhee

The estimation of static parameters in dynamical systems and control theory has been extensively studied, with significant progress made in estimating varying parameters in specific system types. Suppose, in the general case, we have data…

Optimization and Control · Mathematics 2025-07-10 Jamiree Harrison , Enoch Yeung

In this paper, we develop a general law of large numbers and central limit theorem for cumulative reward processes associated with finite state Markov jump processes with non-stationary transition rates. Such models commonly arise in…

Probability · Mathematics 2025-10-15 Monte Fischer , Peter W. Glynn
‹ Prev 1 4 5 6 7 8 10 Next ›