English
Related papers

Related papers: Quadratic Suffices for Over-parametrization via Ma…

200 papers

We consider dynamical and geometrical aspects of deep learning. For many standard choices of layer maps we display semi-invariant metrics which quantify differences between data or decision functions. This allows us, when considering random…

Machine Learning · Computer Science 2021-04-23 Benny Avelin , Anders Karlsson

We design a sublinear-time approximation algorithm for quadratic function minimization problems with a better error bound than the previous algorithm by Hayashi and Yoshida (NIPS'16). Our approximation algorithm can be modified to handle…

Data Structures and Algorithms · Computer Science 2018-06-29 Amit Levi , Yuichi Yoshida

Duan, Wu and Zhou (FOCS 2023) recently obtained the improved upper bound on the exponent of square matrix multiplication $\omega<2.3719$ by introducing a new approach to quantify and compensate the ``combination loss" in prior analyses of…

Data Structures and Algorithms · Computer Science 2023-12-29 François Le Gall

In this note we obtain sharp bounds for the identric mean in terms of a two parameter family of means. Our results generalize and extend recent bounds due to Y. M. Chu & al. (2011), and to M.-K. Wang & al. (2012).

Classical Analysis and ODEs · Mathematics 2018-09-26 Omran Kouba

Following the recent work of Jiang and Lin (Linear Algebra Appl. 585 (2020) 45--49), we present more results (bounds) on Harnack type inequalities for matrices in terms of majorization (i.e., in partial products) of eigenvalues and singular…

Functional Analysis · Mathematics 2019-12-09 Chaojun Yang , Fuzhen Zhang

This note presents absolute bounds on the size of the coefficients of the characteristic and minimal polynomials depending on the size of the coefficients of the associated matrix. Moreover, we present algorithms to compute more precise…

Symbolic Computation · Computer Science 2011-11-10 Jean-Guillaume Dumas

It has been recognized that a heavily overparameterized artificial neural network exhibits surprisingly good generalization performance in various machine-learning tasks. Recent theoretical studies have made attempts to unveil the mystery…

Machine Learning · Computer Science 2021-01-28 Takashi Mori , Masahito Ueda

Neural networks exhibit good generalization behavior in the over-parameterized regime, where the number of network parameters exceeds the number of observations. Nonetheless, current generalization bounds for neural networks fail to explain…

Machine Learning · Computer Science 2017-10-30 Alon Brutzkus , Amir Globerson , Eran Malach , Shai Shalev-Shwartz

We prove new parameterization theorems for sets definable in the structure $\mathbb{R}_{an}$ (i.e. for globally subanalytic sets) which are uniform for definable families of such sets. We treat both $C^r$-parameterization and (mild)…

Number Theory · Mathematics 2018-05-17 Raf Cluckers , Jonathan Pila , Alex Wilkie

We present new results on Boolean matrix factorization and a new algorithm based on these results. The results emphasize the significance of factorizations that provide from-below approximations of the input matrix. While the previously…

Numerical Analysis · Computer Science 2015-06-26 Radim Belohlavek , Martin Trnecka

Modern Machine Learning (ML) and Deep Neural Networks (DNNs) often operate on high-dimensional data and rely on overparameterized models, where classical low-dimensional intuitions break down. In particular, the proportional regime where…

Machine Learning · Statistics 2026-04-17 Zhenyu Liao , Michael W. Mahoney

Deep networks are able to learn highly predictive models of video data. Due to video length, a common strategy is to train them on small video snippets. We apply the deep Taylor / LRP technique to understand the deep network's…

Machine Learning · Computer Science 2018-06-20 Christopher Anders , Grégoire Montavon , Wojciech Samek , Klaus-Robert Müller

This paper addresses the problem of nearly optimal Vapnik--Chervonenkis dimension (VC-dimension) and pseudo-dimension estimations of the derivative functions of deep neural networks (DNNs). Two important applications of these estimations…

Machine Learning · Computer Science 2023-05-16 Yahong Yang , Haizhao Yang , Yang Xiang

Deep learning has demonstrated the power of detailed modeling of complex high-order (multivariate) interactions in data. For some learning tasks there is power in learning models that are not only Deep but also Broad. By Broad, we mean…

Machine Learning · Computer Science 2015-09-07 Nayyar A. Zaidi , Geoffrey I. Webb , Mark J. Carman , Francois Petitjean

Variational mean field approximations tend to struggle with contemporary overparametrized deep neural networks. Where a Bayesian treatment is usually associated with high-quality predictions and uncertainties, the practical reality has been…

A new error bound for the linear complementarity problem when the matrix involved is a B-matrix is presented, which improves the corresponding result in [C.Q. Li et al., A new error bound for linear complementarity problems for B-matrices.…

Numerical Analysis · Mathematics 2016-10-21 Lei Gao , Chaoqian Li

The technique of $Q$-polinomials are used to derive the $w$- constraints in the two-matrix and Kontsevich-like model at finite $N$. These constraints are closed and form Lie algebra. They are associated with the matrices, $\lambda…

High Energy Physics - Theory · Physics 2007-05-23 N. L. Khviengia

The knowledge that data lies close to a particular submanifold of the ambient Euclidean space may be useful in a number of ways. For instance, one may want to automatically mark any point far away from the submanifold as an outlier or to…

The matroid parity (or matroid matching) problem, introduced as a common generalization of matching and matroid intersection problems, is so general that it requires an exponential number of oracle calls. Nevertheless, Lov\'asz (1980)…

Data Structures and Algorithms · Computer Science 2019-06-03 Satoru Iwata , Yusuke Kobayashi

For almost 70 years, researchers have typically selected the width of neural networks' layers either manually or through automated hyperparameter tuning methods such as grid search and, more recently, neural architecture search. This paper…

Machine Learning · Computer Science 2026-02-17 Federico Errica , Henrik Christiansen , Viktor Zaverkin , Mathias Niepert , Francesco Alesiani
‹ Prev 1 8 9 10 Next ›