Related papers: Large deviation principles for renewal-reward proc…
Process Reward Models (PRMs) have emerged as a promising approach to enhance the reasoning capabilities of large language models (LLMs) by guiding their step-by-step reasoning toward a final answer. However, existing PRMs either treat each…
We study large and moderate deviations for a life insurance portfolio, without assuming identically distributed losses. The crucial assumption is that losses are bounded, and that variances are bounded below. From a standard large…
Large language models (LLMs) inevitably make mistakes when performing step-by-step mathematical reasoning. Process Reward Models (PRMs) have emerged as a promising solution by evaluating each reasoning step. However, existing PRMs typically…
This paper considers the Cram\'er-Lundberg model, with the additional feature that the number of clients can fluctuate over time. Clients arrive according to a Poisson process, where the times they spend in the system form a sequence of…
We obtain sharp large deviation estimates for exceedance probabilities in dependent triangular array threshold models with a diverging number of latent factors. The prefactors quantify how latent-factor dependence and tail geometry enter at…
A basic result of large deviations theory is Sanov's theorem, which states that the sequence of empirical measures of independent and identically distributed samples satisfies the large deviation principle with rate function given by…
We show a deviation inequality for U-statistics of independent data taking values in a separable Banach space which satisfies some smoothness assumptions. We then provide applications to rates in the law of large numbers for U-statistics, a…
We study a large deviation principle for a system of stochastic reaction--diffusion equations (SRDEs) with a separation of fast and slow components and small noise in the slow component. The derivation of the large deviation principle is…
We prove a large deviation principle for the sequence of push-forwards of empirical measures in the setting of Riesz potential interactions on compact subsets K in R^d with continuous external fields. Our results are valid for base measures…
Reward models play a critical role in guiding large language models toward outputs that align with human expectations. However, an open challenge remains in effectively utilizing test-time compute to enhance reward model performance. In…
We study the dynamics of smooth interval maps with non-flat critical points. For every such a map that is topologically exact, we establish the full (level-2) Large Deviation Principle for empirical means. In particular, the Large Deviation…
We consider large random trees under Gibbs distributions and prove a Large Deviation Principle (LDP) for the distribution of degrees of vertices of the tree. The LDP rate function is given explicitly. An immediate consequence is a Law of…
This article concerns the large deviations regime and the consequent solution of the Kramers problem for a two-time scale stochastic system driven by a common jump noise signal perturbed in small intensity $\varepsilon>0$ and with…
We establish a large deviation principle for time dependent trajectories (paths) of the empirical density of $N$ particles with long range interactions, for homogeneous systems. This result extends the classical kinetic theory that leads to…
We prove a large deviation principle for the point process of large Poisson $k$-nearest neighbor balls in hyperbolic space. More precisely, we consider a stationary Poisson point process of unit intensity in a growing sampling window in…
We study the cubic weakly nonlinear Schr\"odinger equation with randomized spatially quasi-periodic initial data in higher dimensions. Under a polynomial decay assumption in Fourier space, we establish a {\em Large Deviations Principle} for…
We are dealing with the validity of a large deviation principle for a class of reaction-diffusion equations with polynomial nonlinearity, perturbed by a Gaussian random forcing. We are here interested in the regime where both the strength…
We use a weak Gibbs property and a weak form of specification to derive level-2 large deviations principles for symbolic systems equipped with a large class of reference measures. This has applications to a broad class of symbolic systems,…
We study two problems. First, we consider the large deviation behavior of empirical measures of certain diffusion processes as, simultaneously, the time horizon becomes large and noise becomes vanishingly small. The law of large numbers…
Recent alignment techniques, such as reinforcement learning from human feedback, have been widely adopted to align large language models with human preferences by learning and leveraging reward models. In practice, these models often…