English
Related papers

Related papers: A Gaussian Comparison Theorem for Training Dynamic…

200 papers

Understanding the dynamics of neural network parameters during training is one of the key challenges in building a theoretical foundation for deep learning. A central obstacle is that the motion of a network in high-dimensional parameter…

Machine Learning · Computer Science 2021-03-30 Daniel Kunin , Javier Sagastuy-Brena , Surya Ganguli , Daniel L. K. Yamins , Hidenori Tanaka

We investigate the generalization ability of a perceptron with non-monotonic transfer function of a reversed-wedge type in on-line mode. This network is identical to a parity machine, a multilayer network. We consider several learning…

Disordered Systems and Neural Networks · Physics 2009-10-30 Jun-ichi Inoue , Hidetoshi Nishimori , Yoshiyuki Kabashima

We study estimation and prediction of Gaussian random fields with covariance models belonging to the generalized Wendland (GW) class, under fixed domain asymptotics. As the Mat\'ern case, this class allows a continuous parameterization of…

Statistics Theory · Mathematics 2017-11-17 M. Bevilacqua , T. Faouzi , R. Furrer , E. Porcu

We study learning algorithms when there is a mismatch between the distributions of the training and test datasets of a learning algorithm. The effect of this mismatch on the generalization error and model misspecification are quantified.…

Information Theory · Computer Science 2022-08-11 Saeed Masiha , Amin Gohari , Mohammad Hossein Yassaee , Mohammad Reza Aref

Symmetries play a conspicuous role in the large-scale behavior of critical systems. While in equilibrium they allow to classify asymptotics into different universality classes, out of equilibrium they can emerge, some times unexpectedly, as…

Statistical Mechanics · Physics 2019-04-26 Enrique Rodriguez-Fernandez , Rodolfo Cuerno

We consider the estimation of an n-dimensional vector s from the noisy element-wise measurements of $\mathbf{s}\mathbf{s}^T$, a generic problem that arises in statistics and machine learning. We study a mismatched Bayesian inference…

Information Theory · Computer Science 2021-09-14 Farzad Pourkamali , Nicolas Macris

The softmax cross-entropy loss function has been widely used to train deep models for various tasks. In this work, we propose a Gaussian mixture (GM) loss function for deep neural networks for visual classification. Unlike the softmax…

Computer Vision and Pattern Recognition · Computer Science 2020-11-19 Weitao Wan , Jiansheng Chen , Cheng Yu , Tong Wu , Yuanyi Zhong , Ming-Hsuan Yang

Machine learning systems often acquire biases by leveraging undesired features in the data, impacting accuracy variably across different sub-populations. Current understanding of bias formation mostly focuses on the initial and final stages…

Machine Learning · Computer Science 2024-12-24 Anchit Jain , Rozhin Nobahari , Aristide Baratin , Stefano Sarao Mannelli

We develop an approach for Bayesian learning of spatiotemporal dynamical mechanistic models. Such learning consists of statistical emulation of the mechanistic system that can efficiently interpolate the output of the system from arbitrary…

Methodology · Statistics 2025-07-11 Sudipto Banerjee , Xiang Chen , Ian Frankenburg , Daniel Zhou

Feature selection can facilitate the learning of mixtures of discrete random variables as they arise, e.g. in crowdsourcing tasks. Intuitively, not all workers are equally reliable but, if the less reliable ones could be eliminated, then…

Machine Learning · Statistics 2017-11-28 Vincent Zhao , Steven W. Zucker

Gaussian processes (GPs) are commonplace in spatial statistics. Although many non-stationary models have been developed, there is arguably a lack of flexibility compared to equipping each location with its own parameters. However, the…

Machine Learning · Statistics 2018-07-19 Leo L. Duan , Xia Wang , Rhonda D. Szczesniak

We examine a class of stochastic differential inclusions involving multiscale effects designed to solve a class of generalized variational inequalities. This class of problems contains constrained convex non-smooth optimization problems,…

Optimization and Control · Mathematics 2026-01-23 D. Russell Luke , Johannes-Carl Schnebel , Mathias Staudigl , Juan Peypouquet , Siqi Qu

In this paper, we investigate a general class of stochastic gradient descent (SGD) algorithms, called Conditioned SGD, based on a preconditioning of the gradient direction. Using a discrete-time approach with martingale tools, we establish…

Statistics Theory · Mathematics 2023-10-17 Rémi Leluc , François Portier

Conventional uncertainty-aware temporal difference (TD) learning often assumes a zero-mean Gaussian distribution for TD errors, leading to inaccurate error representations and compromised uncertainty estimation. We introduce a novel…

Machine Learning · Computer Science 2025-02-04 Seyeon Kim , Joonhun Lee , Namhoon Cho , Sungjun Han , Wooseop Hwang

We develop a systematic projection-operator technique for constructing Gaussian approximations and their perturbative corrections in bosonic nonlinear models. As a case study, we apply it to the driven dissipative Kerr oscillator. In the…

Quantum Physics · Physics 2026-03-02 K. Sh. Meretukov , A. E. Teretenkov

The mean-field theory of Kinetically-Constrained-Models is developed by considering the Fredrickson-Andersen model on the Bethe lattice. Using certain properties of the dynamics observed in actual numerical experiments we derive asymptotic…

Disordered Systems and Neural Networks · Physics 2025-01-20 Gianmarco Perrupato , Tommaso Rizzo

Adversarial Imitation Learning is traditionally framed as a two-player zero-sum game between a learner and an adversarially chosen cost function, and can therefore be thought of as the sequential generalization of a Generative Adversarial…

Machine Learning · Computer Science 2025-03-04 Runzhe Wu , Yiding Chen , Gokul Swamy , Kianté Brantley , Wen Sun

Transfer in reinforcement learning is usually achieved through generalisation across tasks. Whilst many studies have investigated transferring knowledge when the reward function changes, they have assumed that the dynamics of the…

Machine Learning · Computer Science 2021-07-20 Majid Abdolshah , Hung Le , Thommen Karimpanal George , Sunil Gupta , Santu Rana , Svetha Venkatesh

We consider the estimation of parametric fractional time series models in which not only is the memory parameter unknown, but one may not know whether it lies in the stationary/invertible region or the nonstationary or noninvertible…

Statistics Theory · Mathematics 2012-03-14 Javier Hualde , Peter M. Robinson

This paper presents a study of the effectiveness of Neural Network (NN) techniques for deconvolution inverse problems relevant for applications in Quantum Field Theory, but also in more general contexts. We consider NN's asymptotic limits,…

Computational Physics · Physics 2024-02-16 Luigi Del Debbio , Manuel Naviglio , Francesco Tarantelli