中文
相关论文

相关论文: Regularization, sparse recovery, and median-of-mea…

200 篇论文

We consider the classical statistical learning/regression problem, when the value of a real random variable Y is to be predicted based on the observation of another random variable X. Given a class of functions F and a sample of independent…

统计理论 · 数学 2016-08-03 Gabor Lugosi , Shahar Mendelson

The goal of regression and classification methods in supervised learning is to minimize the empirical risk, that is, the expectation of some loss function quantifying the prediction error under the empirical distribution. When facing scarce…

最优化与控制 · 数学 2019-07-15 Soroosh Shafieezadeh-Abadeh , Daniel Kuhn , Peyman Mohajerin Esfahani

Tournament procedures, recently introduced in Lugosi & Mendelson (2016), offer an appealing alternative, from a theoretical perspective at least, to the principle of Empirical Risk Minimization in machine learning. Statistical learning by…

机器学习 · 统计学 2022-11-02 Pierre Laforgue , Stephan Clémençon , Patrice Bertail

We survey some of the recent advances in mean estimation and regression function estimation. In particular, we describe sub-Gaussian mean estimators for possibly heavy-tailed data both in the univariate and multivariate settings. We focus…

统计理论 · 数学 2019-06-12 Gabor Lugosi , Shahar Mendelson

Within the statistical and machine learning literature, regularization techniques are often used to construct sparse (predictive) models. Most regularization strategies only work for data where all predictors are treated identically, such…

统计计算 · 统计学 2020-12-16 Sander Devriendt , Katrien Antonio , Tom Reynkens , Roel Verbelen

We develop a novel procedure for estimating the optimizer of general convex stochastic optimization problems of the form $\min_{x\in\mathcal{X}} \mathbb{E}[F(x,\xi)]$, when the given data is a finite independent sample selected according to…

统计理论 · 数学 2022-01-26 Daniel Bartl , Shahar Mendelson

We explore the recent results announced in "Robust machine learning by median-of-means: theory and practice" by G. Lecu\'e and M. Lerasle. We show that these results are, in fact, almost obvious outcomes of the machinery developed in [4]…

统计理论 · 数学 2017-12-20 Gabor Lugosi , Shahar Mendelson

The purpose of this paper is to discuss empirical risk minimization when the losses are not necessarily bounded and may have a distribution with heavy tails. In such situations, usual empirical averages may fail to provide reliable…

统计方法学 · 统计学 2016-08-11 Christian Brownlees , Emilien Joly , Gábor Lugosi

The regsem package in R, an implementation of regularized structural equation modeling (RegSEM; Jacobucci, Grimm, and McArdle 2016), was recently developed with the goal of incorporating various forms of penalized likelihood estimation in a…

统计方法学 · 统计学 2017-09-11 Ross Jacobucci

Many applied settings in empirical economics involve simultaneous estimation of a large number of parameters. In particular, applied economists are often interested in estimating the effects of many-valued treatments (like teacher effects…

机器学习 · 统计学 2017-04-03 Alberto Abadie , Maximilian Kasy

Modern applications require methods that are computationally feasible on large datasets but also preserve statistical efficiency. Frequently, these two concerns are seen as contradictory: approximation methods that enable computation are…

统计方法学 · 统计学 2021-06-11 Darren Homrighausen , Daniel J. McDonald

High-dimensional regression often suffers from heavy-tailed noise and outliers, which can severely undermine the reliability of least-squares based methods. To improve robustness, we adopt a non-smooth Wilcoxon score based rank objective…

机器学习 · 统计学 2026-01-29 Meixia Lin , Meijiao Shi , Yunhai Xiao , Qian Zhang

Unraveling the reasons behind the remarkable success and exceptional generalization capabilities of deep neural networks presents a formidable challenge. Recent insights from random matrix theory, specifically those concerning the spectral…

机器学习 · 统计学 2023-04-10 Xuanzhe Xiao , Zeng Li , Chuanlong Xie , Fengwei Zhou

The paper introduces structured machine learning regressions for heavy-tailed dependent panel data potentially sampled at different frequencies. We focus on the sparse-group LASSO regularization. This type of regularization can take…

计量经济学 · 经济学 2021-11-23 Andrii Babii , Ryan T. Ball , Eric Ghysels , Jonas Striaukas

Current methods for regularization in machine learning require quite specific model assumptions (e.g. a kernel shape) that are not derived from prior knowledge about the application, but must be imposed merely to make the method work. We…

机器学习 · 统计学 2022-11-01 Matthias Wieler

Regularized regression techniques for linear regression have been created the last few ten years to reduce the flaws of ordinary least squares regression with regard to prediction accuracy. In this paper, new methods for using regularized…

机器学习 · 计算机科学 2013-12-13 Doreswamy , Chanabasayya . M. Vastrad

It is well-known that trimmed sample means are robust against heavy tails and data contamination. This paper analyzes the performance of trimmed means and related methods in two novel contexts. The first one consists of estimating…

统计理论 · 数学 2025-12-03 Roberto I. Oliveira , Lucas Resende

Under losses which are potentially heavy-tailed, we consider the task of minimizing sums of the loss mean and standard deviation, without trying to accurately estimate the variance. By modifying a technique for variance-free robust mean…

机器学习 · 统计学 2024-02-12 Matthew J. Holland

Recently there has been a surge of interest in understanding implicit regularization properties of iterative gradient-based optimization algorithms. In this paper, we study the statistical guarantees on the excess risk achieved by…

机器学习 · 统计学 2020-08-28 Tomas Vaškevičius , Varun Kanade , Patrick Rebeschini

This paper introduces a general regularized thresholded least-square procedure estimating a structured signal $\theta_*\in\mathbb{R}^d$ from the following observations: $y_i = f(\langle\mathbf{x}_i, \theta_*\rangle,…

统计理论 · 数学 2018-04-18 Xiaohan Wei
‹ 上一页 1 2 3 10 下一页 ›