English
Related papers

Related papers: Asymptotic behavior of $\ell_p$-based Laplacian re…

200 papers

We analyze recurrent neural networks with diagonal hidden-to-hidden weight matrices, trained with gradient descent in the supervised learning setting, and prove that gradient descent can achieve optimality \emph{without} massive…

Machine Learning · Computer Science 2024-10-11 Semih Cayci , Atilla Eryilmaz

The smallest eigenvectors of the graph Laplacian are well-known to provide a succinct representation of the geometry of a weighted graph. In reinforcement learning (RL), where the weighted graph may be interpreted as the state transition…

Machine Learning · Computer Science 2018-10-11 Yifan Wu , George Tucker , Ofir Nachum

Regularization plays a pivotal role when facing the challenge of solving ill-posed inverse problems, where the number of observations is smaller than the ambient dimension of the object to be estimated. A line of recent work has studied…

Optimization and Control · Mathematics 2014-07-03 Samuel Vaiter , Mohammad Golbabaee , Jalal M. Fadili , Gabriel Peyré

We study the theory of neural network (NN) from the lens of classical nonparametric regression problems with a focus on NN's ability to adaptively estimate functions with heterogeneous smoothness -- a property of functions in Besov or…

Machine Learning · Computer Science 2024-05-21 Kaiqi Zhang , Yu-Xiang Wang

The thesis studies linear and semilinear Dirichlet problems driven by different fractional Laplacians. The boundary data can be smooth functions or also Radon measures. The goal is to classify the solutions which have a singularity on the…

Analysis of PDEs · Mathematics 2015-11-03 Nicola Abatangelo

We discuss stability for a class of learning algorithms with respect to noisy labels. The algorithms we consider are for regression, and they involve the minimization of regularized risk functionals, such as L(f) := 1/N sum_i…

Machine Learning · Computer Science 2007-05-23 Cynthia Rudin

We provide a statistical analysis of regularization-based continual learning on a sequence of linear regression tasks, with emphasis on how different regularization terms affect the model performance. We first derive the convergence rate…

Machine Learning · Computer Science 2024-06-11 Xuyang Zhao , Huiyuan Wang , Weiran Huang , Wei Lin

Existing approaches to analyzing the asymptotics of graph Laplacians typically assume a well-behaved kernel function with smoothness assumptions. We remove the smoothness assumption and generalize the analysis of graph Laplacians to include…

Machine Learning · Statistics 2011-01-31 Daniel Ting , Ling Huang , Michael Jordan

In this manuscript, we obtain sharp and improved regularity estimates for weak solutions of weighted quasilinear elliptic models of Hardy-H\'{e}non-type, featuring an explicit regularity exponent depending only on universal parameters. Our…

Analysis of PDEs · Mathematics 2024-10-22 João Vitor da Silva , Disson dos Prazeres , Gleydson Ricarte , Ginaldo Sá

In many applications, one has side information, e.g., labels that are provided in a semi-supervised manner, about a specific target region of a large data set, and one wants to perform machine learning and data analysis tasks "nearby" that…

Machine Learning · Computer Science 2013-04-30 Toke J. Hansen , Michael W. Mahoney

Recently, Mahoney and Orecchia demonstrated that popular diffusion-based procedures to compute a quick \emph{approximation} to the first nontrivial eigenvector of a data graph Laplacian \emph{exactly} solve certain regularized Semi-Definite…

Data Structures and Algorithms · Computer Science 2011-10-13 Patrick O. Perry , Michael W. Mahoney

Modern machine learning models are often trained in a setting where the number of parameters exceeds the number of training samples. To understand the implicit bias of gradient descent in such overparameterized models, prior work has…

Machine Learning · Statistics 2025-10-29 Hannes Matt , Dominik Stöger

Graph convolutional networks (GCN) are viewed as one of the most popular representations among the variants of graph neural networks over graph data and have shown powerful performance in empirical experiments. That $\ell_2$-based graph…

Machine Learning · Computer Science 2023-06-21 Shiyu Liu , Linsen Wei , Shaogao Lv , Ming Li

We establish the $L_p$-regularity theory for a semilinear stochastic partial differential equation with multiplicative white noise: $$ du = (a^{ij}u_{x^ix^j} + b^{i}u_{x^i} + cu + \bar b^{i}|u|^\lambda u_{x^i})dt + \sigma^k(u)dw_t^k,\quad…

Probability · Mathematics 2022-05-24 Beom-Seok Han

We systematically explore regularizing neural networks by penalizing low entropy output distributions. We show that penalizing low entropy output distributions, which has been shown to improve exploration in reinforcement learning, acts as…

Neural and Evolutionary Computing · Computer Science 2017-01-24 Gabriel Pereyra , George Tucker , Jan Chorowski , Łukasz Kaiser , Geoffrey Hinton

Clustering (or community detection) on multilayer graphs poses several additional complications with respect to standard graphs as different layers may be characterized by different structures and types of information. One of the major…

Machine Learning · Computer Science 2023-06-02 Sara Venturini , Andrea Cristofari , Francesco Rinaldi , Francesco Tudisco

We study minimax lower bounds for function estimation problems on large graph when the target function is smoothly varying over the graph. We derive minimax rates in the context of regression and classification problems on graphs that…

Statistics Theory · Mathematics 2018-02-16 Alisa Kirichenko , Harry van Zanten

In overparametrized models, the noise in stochastic gradient descent (SGD) implicitly regularizes the optimization trajectory and determines which local minimum SGD converges to. Motivated by empirical studies that demonstrate that training…

Machine Learning · Computer Science 2021-12-07 Alex Damian , Tengyu Ma , Jason D. Lee

We present a novel cost function for semi-supervised learning of neural networks that encourages compact clustering of the latent space to facilitate separation. The key idea is to dynamically create a graph over embeddings of labeled and…

In this paper, we study a nonlocal variational problem which consists of minimizing in $L^2$ the sum of a quadratic data fidelity and a regularization term corresponding to the $L^p$-norm of the nonlocal gradient. In particular, we study…

Numerical Analysis · Mathematics 2019-08-21 Yosra Hafiene , Jalal Fadili , Abderrahim Elmoataz