English
Related papers

Related papers: Lower large deviations for geometric functionals i…

200 papers

We study nonasymptotic minimax estimation of the linear functional $L(\theta)=\eta^\top \theta$ for a high-dimensional $s$-sparse mean vector with an arbitrary loading vector $\eta$. For symmetric noise with exponentially decaying tails, we…

Statistics Theory · Mathematics 2026-04-29 Jie Xie , Dongming Huang

We consider large deviations of empirical measures of diffusion processes. In a first part, we present conditions to obtain a large deviations principle (LDP) for a precise class of unbounded functions. This provides an analogue to the…

Probability · Mathematics 2020-09-23 Grégoire Ferré , Gabriel Stoltz

Random feature methods have been successful in various machine learning tasks, are easy to compute, and come with theoretical accuracy bounds. They serve as an alternative approach to standard neural networks since they can represent…

Machine Learning · Statistics 2026-01-21 Abolfazl Hashemi , Hayden Schaeffer , Robert Shi , Ufuk Topcu , Giang Tran , Rachel Ward

We propose and investigate a unifying class of sparse random graph models, based on a hidden coloring of edge-vertex incidences, extending an existing approach, Random graphs with a given degree distribution, in a way that admits a…

Statistical Mechanics · Physics 2009-11-10 Bo Söderberg

The symmetric simple exclusion process is one of the simplest out-of-equilibrium systems for which the steady state is known. Its large deviation functional of the density has been computed in the past both by microscopic and macroscopic…

Statistical Mechanics · Physics 2015-06-15 B. Derrida , M. Retaux

This article suggests that deterministic Gradient Descent, which does not use any stochastic gradient approximation, can still exhibit stochastic behaviors. In particular, it shows that if the objective function exhibit multiscale…

Machine Learning · Computer Science 2020-11-03 Lingkai Kong , Molei Tao

We consider a random walk in a random environment (RWRE) on the strip of finite width $\mathbb{Z} \times \{1,2,\ldots,d\}$. We prove both quenched and averaged large deviation principles for the position and the hitting times of the RWRE.…

Probability · Mathematics 2016-06-20 Jonathon Peterson

Sharp large deviation estimates for stochastic differential equations with small noise, based on minimizing the Freidlin-Wentzell action functional under appropriate boundary conditions, can be obtained by integrating certain matrix Riccati…

Statistical Mechanics · Physics 2023-01-11 Timo Schorlepp , Tobias Grafke , Rainer Grauer

Classical optimisation theory guarantees monotonic objective decrease for gradient descent (GD) when employed in a small step size, or ``stable", regime. In contrast, gradient descent on neural networks is frequently performed in a large…

Machine Learning · Computer Science 2025-10-21 Lachlan Ewen MacDonald , Hancheng Min , Leandro Palma , Salma Tarmoun , Ziqing Xu , René Vidal

We consider the efficient numerical minimization of Tikhonov functionals with nonlinear operators and non-smooth and non-convex penalty terms, which appear for example in variational regularization. For this, we consider a new class of SCD…

Numerical Analysis · Mathematics 2024-10-18 Helmut Gfrerer , Simon Hubmer , Ronny Ramlau

In this paper, we present and analyze a staggered discontinuous Galerkin method for Darcy flows in fractured porous media on fairly general meshes. A staggered discontinuous Galerkin method and a standard conforming finite element method…

Numerical Analysis · Mathematics 2020-05-25 Lina Zhao , Dohyun Kim , Eun-Jae Park , Eric Chung

In this paper the efficiency of multilevel sparse tensor approximation methods for high-dimensional affine parametric diffusion equations is investigated. Methodologically, the recently presented Sparse Alternating Least Squares (SALS)…

Numerical Analysis · Mathematics 2026-03-17 Martin Eigel , Philipp Trunschke , Dana Wrischnig

Stochastic gradient descent (SGD) is almost ubiquitously used for training non-convex optimization tasks. Recently, a hypothesis proposed by Keskar et al. [2017] that large batch methods tend to converge to sharp minimizers has received…

Machine Learning · Statistics 2018-12-04 Xiaowu Dai , Yuhua Zhu

This paper constitutes our initial effort in developing sparse grid discontinuous Galerkin (DG) methods for high-dimensional partial differential equations (PDEs). Over the past few decades, DG methods have gained popularity in many…

Numerical Analysis · Mathematics 2016-04-20 Zixuan Wang , Qi Tang , Wei Guo , Yingda Cheng

We prove the reduction principle for asymptotics of functionals of vector random fields with weakly and strongly dependent components. These functionals can be used to construct new classes of random fields with skewed and heavy-tailed…

Probability · Mathematics 2020-05-04 Andriy Olenko , Dareen Omari

In this paper, we use the framework of mod-$\phi$ convergence to prove precise large or moderate deviations for quite general sequences of real valued random variables $(X_{n})_{n \in \mathbb{N}}$, which can be lattice or non-lattice…

Probability · Mathematics 2017-02-14 Valentin Féray , Pierre-Loïc Méliot , Ashkan Nikeghbali

Sketched gradient algorithms have been recently introduced for efficiently solving the large-scale constrained Least-squares regressions. In this paper we provide novel convergence analysis for the basic method {\it Gradient Projection…

Optimization and Control · Mathematics 2017-06-05 Junqi Tang , Mohammad Golbabaee , Mike Davies

We investigate nonparametric regression methods based on spatial depth and quantiles when the response and the covariate are both functions. As in classical quantile regression for finite dimensional data, regression techniques developed…

Methodology · Statistics 2018-02-14 Joydeep Chowdhury , Probal Chaudhuri

The paper explores the differential inclusion of a special form. It is supposed that the support function of the set in the right-hand side of an inclusion may contain the maximum of the finite number of continuously differentiable (in…

Optimization and Control · Mathematics 2023-05-04 Alexander Fominyh

In this paper, we want to clarify the Gibbs phenomenon when continuous and discontinuous finite elements are used to approximate discontinuous or nearly discontinuous PDE solutions from the approximation point of view. For a simple step…

Numerical Analysis · Mathematics 2022-08-03 Shun Zhang