English
Related papers

Related papers: Large deviation principles for renewal-reward proc…

200 papers

Process Reward Models (PRMs) have emerged as a promising approach to enhance the reasoning capabilities of large language models (LLMs) by guiding their step-by-step reasoning toward a final answer. However, existing PRMs either treat each…

Machine Learning · Computer Science 2026-03-02 Zheng Zhang , Ziwei Shan , Kaitao Song , Yexin Li , Kan Ren

We study large and moderate deviations for a life insurance portfolio, without assuming identically distributed losses. The crucial assumption is that losses are bounded, and that variances are bounded below. From a standard large…

Probability · Mathematics 2020-09-04 Stefan Gerhold

Large language models (LLMs) inevitably make mistakes when performing step-by-step mathematical reasoning. Process Reward Models (PRMs) have emerged as a promising solution by evaluating each reasoning step. However, existing PRMs typically…

Computation and Language · Computer Science 2025-03-28 Shuaijie She , Junxiao Liu , Yifeng Liu , Jiajun Chen , Xin Huang , Shujian Huang

This paper considers the Cram\'er-Lundberg model, with the additional feature that the number of clients can fluctuate over time. Clients arrive according to a Poisson process, where the times they spend in the system form a sequence of…

Probability · Mathematics 2023-05-25 Peter Braunsteins , Michel Mandjes

We obtain sharp large deviation estimates for exceedance probabilities in dependent triangular array threshold models with a diverging number of latent factors. The prefactors quantify how latent-factor dependence and tail geometry enter at…

Probability · Mathematics 2025-10-21 Fengnan Deng , Anand N. Vidyashankar , Jeffrey F. Collamore

A basic result of large deviations theory is Sanov's theorem, which states that the sequence of empirical measures of independent and identically distributed samples satisfies the large deviation principle with rate function given by…

Probability · Mathematics 2014-10-17 Markus Fischer

We show a deviation inequality for U-statistics of independent data taking values in a separable Banach space which satisfies some smoothness assumptions. We then provide applications to rates in the law of large numbers for U-statistics, a…

Probability · Mathematics 2024-05-06 Davide Giraudo

We study a large deviation principle for a system of stochastic reaction--diffusion equations (SRDEs) with a separation of fast and slow components and small noise in the slow component. The derivation of the large deviation principle is…

Probability · Mathematics 2019-05-02 Wenqing Hu , Michael Salins , Konstantinos Spiliopoulos

We prove a large deviation principle for the sequence of push-forwards of empirical measures in the setting of Riesz potential interactions on compact subsets K in R^d with continuous external fields. Our results are valid for base measures…

Classical Analysis and ODEs · Mathematics 2016-10-27 Tom Bloom , Norman Levenberg , Franck Wielonsky

Reward models play a critical role in guiding large language models toward outputs that align with human expectations. However, an open challenge remains in effectively utilizing test-time compute to enhance reward model performance. In…

Computation and Language · Computer Science 2025-05-21 Jiaxin Guo , Zewen Chi , Li Dong , Qingxiu Dong , Xun Wu , Shaohan Huang , Furu Wei

We study the dynamics of smooth interval maps with non-flat critical points. For every such a map that is topologically exact, we establish the full (level-2) Large Deviation Principle for empirical means. In particular, the Large Deviation…

Dynamical Systems · Mathematics 2019-07-19 Yong Moo Chung , Juan Rivera-Letelier , Hiroki Takahasi

We consider large random trees under Gibbs distributions and prove a Large Deviation Principle (LDP) for the distribution of degrees of vertices of the tree. The LDP rate function is given explicitly. An immediate consequence is a Law of…

Probability · Mathematics 2009-11-13 Yuri Bakhtin , Christine Heitsch

This article concerns the large deviations regime and the consequent solution of the Kramers problem for a two-time scale stochastic system driven by a common jump noise signal perturbed in small intensity $\varepsilon>0$ and with…

Probability · Mathematics 2022-07-15 Pedro Catuogno , André de Oliveira Gomes

We establish a large deviation principle for time dependent trajectories (paths) of the empirical density of $N$ particles with long range interactions, for homogeneous systems. This result extends the classical kinetic theory that leads to…

Statistical Mechanics · Physics 2022-01-19 Ouassim Feliachi , Freddy Bouchet

We prove a large deviation principle for the point process of large Poisson $k$-nearest neighbor balls in hyperbolic space. More precisely, we consider a stationary Poisson point process of unit intensity in a growing sampling window in…

Probability · Mathematics 2023-04-19 Christian Hirsch , Moritz Otto , Takashi Owada , Christoph Thäle

We study the cubic weakly nonlinear Schr\"odinger equation with randomized spatially quasi-periodic initial data in higher dimensions. Under a polynomial decay assumption in Fourier space, we establish a {\em Large Deviations Principle} for…

Probability · Mathematics 2026-04-21 Fei Xu , Yong Li

We are dealing with the validity of a large deviation principle for a class of reaction-diffusion equations with polynomial nonlinearity, perturbed by a Gaussian random forcing. We are here interested in the regime where both the strength…

Probability · Mathematics 2017-05-02 Sandra Cerrai , Arnaud Debussche

We use a weak Gibbs property and a weak form of specification to derive level-2 large deviations principles for symbolic systems equipped with a large class of reference measures. This has applications to a broad class of symbolic systems,…

Dynamical Systems · Mathematics 2017-10-25 Vaughn Climenhaga , Daniel J. Thompson , Kenichiro Yamamoto

We study two problems. First, we consider the large deviation behavior of empirical measures of certain diffusion processes as, simultaneously, the time horizon becomes large and noise becomes vanishingly small. The law of large numbers…

Probability · Mathematics 2023-09-14 Amarjit Budhiraja , Pavlos Zoubouloglou

Recent alignment techniques, such as reinforcement learning from human feedback, have been widely adopted to align large language models with human preferences by learning and leveraging reward models. In practice, these models often…

Machine Learning · Computer Science 2025-10-29 Ignavier Ng , Patrick Blöbaum , Siddharth Bhandari , Kun Zhang , Shiva Kasiviswanathan