中文
相关论文

相关论文: Agnostic Sample Compression Schemes for Regression

200 篇论文

We aim at estimating a function $\lambda:[0,1]\to \mathbb {R}$, subject to the constraint that it is decreasing (or increasing). We provide a unified approach for studying the $\mathbb {L}_p$-loss of an estimator defined as the slope of a…

统计理论 · 数学 2009-09-29 Cécile Durot

We consider $L^2$-approximation on weighted reproducing kernel Hilbert spaces of functions depending on infinitely many variables. We focus on unrestricted linear information, admitting evaluations of arbitrary continuous linear…

数值分析 · 数学 2026-01-13 Kumar Harsha , Michael Gnewuch , Marcin Wnuk

In this paper, we study the support recovery guarantees of underdetermined sparse regression using the $\ell_1$-norm as a regularizer and a non-smooth loss function for data fidelity. More precisely, we focus in detail on the cases of…

信息论 · 计算机科学 2016-11-04 Kévin Degraux , Gabriel Peyré , Jalal M. Fadili , Laurent Jacques

For Euclidean space ($\ell_2$), there exists the powerful dimension reduction transform of Johnson and Lindenstrauss, with a host of known applications. Here, we consider the problem of dimension reduction for all $\ell_p$ spaces $1 \le p…

计算几何 · 计算机科学 2015-12-08 Yair Bartal , Lee-Ad Gottlieb

There are many settings where researchers are interested in estimating average treatment effects and are willing to rely on the unconfoundedness assumption, which requires that the treatment assignment be as good as random conditional on…

统计方法学 · 统计学 2018-02-02 Susan Athey , Guido W. Imbens , Stefan Wager

The goal of compressed sensing is to estimate a vector from an underdetermined system of noisy linear measurements, by making use of prior knowledge on the structure of vectors in the relevant domain. For almost all results in this…

机器学习 · 统计学 2017-03-10 Ashish Bora , Ajil Jalal , Eric Price , Alexandros G. Dimakis

This paper studies the asymptotics of resampling without replacement in the proportional regime where dimension $p$ and sample size $n$ are of the same order. For a given dataset $(X,y)\in \mathbb{R}^{n\times p}\times \mathbb{R}^n$ and…

统计理论 · 数学 2026-02-04 Pierre C. Bellec , Takuya Koriyama

We consider both $\ell _{0}$-penalized and $\ell _{0}$-constrained quantile regression estimators. For the $\ell _{0}$-penalized estimator, we derive an exponential inequality on the tail probability of excess quantile prediction risk and…

统计方法学 · 统计学 2023-03-30 Le-Yu Chen , Sokbae Lee

Sparse linear regression with ill-conditioned Gaussian random designs is widely believed to exhibit a statistical/computational gap, but there is surprisingly little formal evidence for this belief, even in the form of examples that are…

数据结构与算法 · 计算机科学 2022-03-08 Jonathan A. Kelner , Frederic Koehler , Raghu Meka , Dhruv Rohatgi

This paper develops inferential methods for a very general class of ill-posed models in econometrics encompassing the nonparametric instrumental variable regression, various functional regressions, and the density deconvolution. We focus on…

统计理论 · 数学 2020-12-22 Andrii Babii

Sparse recovery is among the most well-studied problems in learning theory and high-dimensional statistics. In this work, we investigate the statistical and computational landscapes of sparse recovery with $\ell_\infty$ error guarantees.…

统计理论 · 数学 2026-02-19 Ziyun Chen , Jerry Li , Kevin Tian , Yusong Zhu

Existing theories on deep nonparametric regression have shown that when the input data lie on a low-dimensional manifold, deep neural networks can adapt to the intrinsic data structures. In real world applications, such an assumption of…

机器学习 · 计算机科学 2023-06-27 Zixuan Zhang , Minshuo Chen , Mengdi Wang , Wenjing Liao , Tuo Zhao

Large language models (LLMs) have shown remarkable success in language modelling due to scaling laws found in model size and the hidden dimension of the model's text representation. Yet, we demonstrate that compressed representations of…

计算与语言 · 计算机科学 2025-02-05 Felix Drinkall , Janet B. Pierrehumbert , Stefan Zohren

Mixed linear regression is a well-studied problem in parametric statistics and machine learning. Given a set of samples, tuples of covariates and labels, the task of mixed linear regression is to find a small list of linear relationships…

机器学习 · 统计学 2024-06-04 Avishek Ghosh , Arya Mazumdar

In our companion work \cite{Stojnicl1RegPosasymldp} we revisited random under-determined linear systems with sparse solutions. The main emphasis was on the performance analysis of the $\ell_1$ heuristic in the so-called asymptotic regime,…

最优化与控制 · 数学 2016-12-20 Mihailo Stojnic

Recently a number of CNN-based techniques were proposed to remove image compression artifacts. As in other restoration applications, these techniques all learn a mapping from decompressed patches to the original counterparts under the…

计算机视觉与模式识别 · 计算机科学 2020-01-22 Xi Zhang , Xiaolin Wu

Consider the use of $\ell_{1}/\ell_{\infty}$-regularized regression for joint estimation of a $\pdim \times \numreg$ matrix of regression coefficients. We analyze the high-dimensional scaling of $\ell_1/\ell_\infty$-regularized quadratic…

统计理论 · 数学 2009-05-12 S. Negahban , M. J. Wainwright

We examine the rate of convergence of the Lasso estimator of lower dimensional components of the high-dimensional parameter. Under bounds on the $\ell_1$-norm on the worst possible sub-direction these rates are of order $\sqrt {|J| \log p /…

统计理论 · 数学 2014-03-28 Sara van de Geer

We consider optimal non-sequential designs for a large class of (linear and nonlinear) regression models involving polynomials and rational functions with heteroscedastic noise also given by a polynomial or rational weight function. The…

统计计算 · 统计学 2011-08-30 Dávid Papp

The adoption of Foundation Models in resource-constrained environments remains challenging due to their large size and inference costs. A promising way to overcome these limitations is post-training compression, which aims to balance…