English
Related papers

Related papers: An Inequality with Applications to Structured Spar…

200 papers

A wide variety of problems in machine learning, including exemplar clustering, document summarization, and sensor placement, can be cast as constrained submodular maximization problems. Unfortunately, the resulting submodular optimization…

Machine Learning · Computer Science 2015-04-23 Rafael da Ponte Barbosa , Alina Ene , Huy L. Nguyen , Justin Ward

We present an extension of sparse PCA, or sparse dictionary learning, where the sparsity patterns of all dictionary elements are structured and constrained to belong to a prespecified set of shapes. This \emph{structured sparse PCA} is…

Machine Learning · Statistics 2009-09-09 Rodolphe Jenatton , Guillaume Obozinski , Francis Bach

We discuss applications of some concepts of Compressed Sensing in the recent work on invertibility of random matrices due to Rudelson and the author. We sketch an argument leading to the optimal bound N^{-1/2} on the median of the smallest…

Numerical Analysis · Mathematics 2016-12-23 Roman Vershynin

Converting an n-dimensional vector to a probability distribution over n objects is a commonly used component in many machine learning tasks like multiclass classification, multilabel classification, attention mechanisms etc. For this,…

Topic sparsity refers to the observation that individual documents usually focus on several salient topics instead of covering a wide variety of topics, and a real topic adopts a narrow range of terms instead of a wide coverage of the…

Machine Learning · Computer Science 2018-11-29 Tianyi Lin , Zhiyue Hu , Xin Guo

Multivariate categorical data occur in many applications of machine learning. One of the main difficulties with these vectors of categorical variables is sparsity. The number of possible observations grows exponentially with vector length,…

Machine Learning · Statistics 2015-03-10 Yarin Gal , Yutian Chen , Zoubin Ghahramani

Stochastic iterative methods are useful in a variety of large-scale numerical linear algebraic, machine learning, and statistical problems, in part due to their low-memory footprint. They are frequently used in a variety of applications,…

Numerical Analysis · Mathematics 2025-11-27 Toby Anderson , Max Collins , Jamie Haddock , Jackie Lok , Elizaveta Rebrova

There has been increasing interest on summary-free solutions for approximate Bayesian computation (ABC) which replace distances among summaries with discrepancies between the empirical distributions of the observed data and the synthetic…

Methodology · Statistics 2025-01-27 Sirio Legramanti , Daniele Durante , Pierre Alquier

Sparse methods for supervised learning aim at finding good linear predictors from as few variables as possible, i.e., with small cardinality of their supports. This combinatorial selection problem is often turned into a convex optimization…

Machine Learning · Computer Science 2010-11-15 Francis Bach

Dictionary learning algorithms have been successfully used for both reconstructive and discriminative tasks, where an input signal is represented with a sparse linear combination of dictionary atoms. While these methods are mostly developed…

Machine Learning · Statistics 2016-01-20 Soheil Bahrampour , Nasser M. Nasrabadi , Asok Ray , W. Kenneth Jenkins

Via a covariance representation based on characteristic functions, a known elementary proof of the Gaussian concentration inequality is presented. A few other applications are briefly mentioned.

Probability · Mathematics 2024-10-10 Christian Houdré

We prove an extension of McDiarmid's inequality for metric spaces with unbounded diameter. To this end, we introduce the notion of the {\em subgaussian diameter}, which is a distribution-dependent refinement of the metric diameter. Our…

Probability · Mathematics 2013-09-12 Aryeh Kontorovich

Sparse coding or sparse dictionary learning has been widely used to recover underlying structure in many kinds of natural data. Here, we provide conditions guaranteeing when this recovery is universal; that is, when sparse codes and…

Neurons and Cognition · Quantitative Biology 2016-11-18 Christopher J. Hillar , Friedrich T. Sommer

Optimization of complex functions, such as the output of computer simulators, is a difficult task that has received much attention in the literature. A less studied problem is that of optimization under unknown constraints, i.e., when the…

Methodology · Statistics 2010-07-06 Robert B. Gramacy , Herbert K. H. Lee

We prove logarithmic Sobolev inequalities and concentration results for convex functions and a class of product random vectors. The results are used to derive tail and moment inequalities for chaos variables (in spirit of Talagrand and…

Probability · Mathematics 2007-05-23 Radoslaw Adamczak

Approximating adequate number of clusters in multidimensional data is an open area of research, given a level of compromise made on the quality of acceptable results. The manuscript addresses the issue by formulating a transductive…

Computer Vision and Pattern Recognition · Computer Science 2015-03-17 Shriprakash Sinha

The main objective of this paper is to prove a new inequality for plurisubharmonic functions estimating their supremum over a ball by their supremum over a measurable subset of the ball. We apply this result to study local properties of…

Complex Variables · Mathematics 2016-09-07 Alexander Brudnyi

In dictionary learning, also known as sparse coding, the algorithm is given samples of the form $y = Ax$ where $x\in \mathbb{R}^m$ is an unknown random sparse vector and $A$ is an unknown dictionary matrix in $\mathbb{R}^{n\times m}$…

Data Structures and Algorithms · Computer Science 2014-01-06 Sanjeev Arora , Aditya Bhaskara , Rong Ge , Tengyu Ma

Statistical learning theory provides bounds of the generalization gap, using in particular the Vapnik-Chervonenkis dimension and the Rademacher complexity. An alternative approach, mainly studied in the statistical physics literature, is…

Disordered Systems and Neural Networks · Physics 2020-09-04 Alia Abbara , Benjamin Aubin , Florent Krzakala , Lenka Zdeborová

We derive sharp upper and lower bounds for the pointwise concentration function of the maximum statistic of $d$ identically distributed real-valued random variables. Our first main result places no restrictions either on the common marginal…

Statistics Theory · Mathematics 2025-08-04 Matias D. Cattaneo , Ricardo P. Masini , William G. Underwood