English
Related papers

Related papers: Double Descent Risk and Volume Saturation Effects:…

200 papers

This manuscript studies statistical properties of linear classifiers obtained through minimization of an unregularized convex risk over a finite sample. Although the results are explicitly finite-dimensional, inputs may be passed through…

Machine Learning · Computer Science 2012-06-15 Matus Telgarsky

Recently, the benefit of heavily overparameterized models has been observed in machine learning tasks: models with enough capacity to easily cross the \emph{interpolation threshold} improve in generalization error compared to the classical…

High Energy Physics - Experiment · Physics 2025-09-03 Matthias Vigl , Lukas Heinrich

We apply random matrix theory to study the impact of measurement uncertainty on dynamic mode decomposition. Specifically, when the measurements follow a normal probability density function, we show how the moments of that density propagate…

Methodology · Statistics 2025-09-04 P. Algikar , P. Sharma , M. Netto , L. Mili

This paper presents a unified framework, integrating information theory and statistical mechanics, to connect metric failure in high-dimensional data with emergence in complex systems. We propose the "Information Dilution Theorem,"…

Information Theory · Computer Science 2025-04-15 HongZheng Liu , YiNuo Tian , Zhiyue Wu

This paper investigates the effect of initial volume fraction on the runout characteristics of collapse of granular columns on slopes in fluid. Two-dimensional sub-grain scale numerical simulations are performed to understand the flow…

Geophysics · Physics 2017-06-30 Krishna Kumar , Jean-Yves Delenne , Kenichi Soga

Continual learning, focused on sequentially learning multiple tasks, has gained significant attention recently. Despite the tremendous progress made in the past, the theoretical understanding, especially factors contributing to catastrophic…

Machine Learning · Computer Science 2024-05-29 Meng Ding , Kaiyi Ji , Di Wang , Jinhui Xu

We develop a technique using dual mixed-volumes to study the isotropic constants of some classes of spaces. In particular, we recover, strengthen and generalize results of Ball and Junge concerning the isotropic constants of subspaces and…

Functional Analysis · Mathematics 2007-05-23 Emanuel Milman

We study the closure properties of the class of Bivariate Regular Variation, symbolically BRV , in standard and nonstandard cases, with respect to the randomly weighted sums. However, we take into consideration a weak dependence structure…

Probability · Mathematics 2025-06-24 Dimitrios G. Konstantinides , Charalampos D. Passalidis

This paper advances the understanding of how the size of a machine learning model affects its vulnerability to poisoning, despite state-of-the-art defenses. Given isotropic random honest feature vectors and the geometric median (or clipped…

Machine Learning · Computer Science 2024-09-27 Lê-Nguyên Hoang

A quadratic approximation of neural network loss landscapes has been extensively used to study the optimization process of these networks. Though, it usually holds in a very small neighborhood of the minimum, it cannot explain many…

Machine Learning · Computer Science 2022-06-23 Chao Ma , Daniel Kunin , Lei Wu , Lexing Ying

Random operators constitute fundamental building blocks of models of complex systems yet are far from fully understood. Here, we explain an asymmetry emerging upon repeating identical isotropic (uniformly random) operations. Specifically,…

Statistical Mechanics · Physics 2021-06-03 Malte Schröder , Marc Timme

Grokking, the phenomenon of delayed generalization, is often attributed to the depth and compositional structure of deep neural networks. We study grokking in one of the simplest possible settings: the learning of a linear model with…

Machine Learning · Computer Science 2026-02-10 Nataraj Das , Atreya Vedantam , Chandrashekar Lakshminarayanan

Recent work by Woodworth et al. (2020) shows that the optimization dynamics of gradient descent for overparameterized problems can be viewed as low-dimensional dual dynamics induced by a mirror map, explaining the implicit regularization…

Machine Learning · Computer Science 2024-10-21 Shuyang Wang , Diego Klabjan

We discuss the systematic effects arising from the cosmological redshift-space (geometric) distortion on the statistical analysis of isodensity contour using high-redshift catalogs. Especially, we present a simple theoretical model for…

Astrophysics · Physics 2009-10-31 Atsushi Taruya , Kazuhiro Yamamoto

Fitting a function by using linear combinations of a large number $N$ of `simple' components is one of the most fruitful ideas in statistical learning. This idea lies at the core of a variety of methods, from two-layer neural networks to…

Statistics Theory · Mathematics 2019-08-20 Adel Javanmard , Marco Mondelli , Andrea Montanari

Cross-validation techniques for risk estimation and model selection are widely used in statistics and machine learning. However, the understanding of the theoretical properties of learning via model selection with cross-validation risk…

Machine Learning · Statistics 2024-05-27 Diego Marcondes , Cláudia Peixoto

The relationship between the number of training data points, the number of parameters, and the generalization capabilities of models has been widely studied. Previous work has shown that double descent can occur in the over-parameterized…

Machine Learning · Statistics 2024-10-28 Xinyue Li , Rishi Sonthalia

Understanding the inductive bias and generalization properties of large overparametrized machine learning models requires to characterize the dynamics of the training algorithm. We study the learning dynamics of large two-layer neural…

Machine Learning · Statistics 2025-10-30 Andrea Montanari , Pierfrancesco Urbani

Bayesian model selection commonly relies on Laplace approximation or the Bayesian Information Criterion (BIC), which assume that the effective model dimension equals the number of parameters. Singular learning theory replaces this…

Machine Learning · Statistics 2026-01-06 Kalyaan Rao

Understanding how the test risk scales with model complexity is a central question in machine learning. Classical theory is challenged by the learning curves observed for large over-parametrized deep networks. Capacity measures based on…

Machine Learning · Statistics 2025-10-22 Yichen Wang , Yudong Chen , Lorenzo Rosasco , Fanghui Liu
‹ Prev 1 4 5 6 7 8 10 Next ›