中文
相关论文

相关论文: Minimax Lower Bounds for Ridge Combinations Includ…

200 篇论文

We propose a nested reduced-rank regression (NRRR) approach in fitting regression model with multivariate functional responses and predictors, to achieve tailored dimension reduction and facilitate interpretation/visualization of the…

统计方法学 · 统计学 2020-03-11 Xiaokang Liu , Shujie Ma , Kun Chen

We consider additive functionals $X_n(\phi)$ with small toll functions on split trees and a generalization of split trees, which we call fractional split trees, where the split vector does not need to sum up to 1. These additive functionals…

概率论 · 数学 2026-03-24 Cecilia Holmgren , Jasper Ischebeck , Svante Janson

Ridge regression (RR) is a regularization technique that penalizes the L2-norm of the coefficients in linear regression. One of the challenges of using RR is the need to set a hyperparameter ($\alpha$) that controls the amount of…

统计方法学 · 统计学 2020-05-08 Ariel Rokem , Kendrick Kay

It has been experimentally observed in recent years that multi-layer artificial neural networks have a surprising ability to generalize, even when trained with far more parameters than observations. Is there a theoretical basis for this?…

机器学习 · 统计学 2018-09-19 Andrew R. Barron , Jason M. Klusowski

In this paper we study the kernel multiple ridge regression framework, which we refer to as multi-task regression, using penalization techniques. The theoretical analysis of this problem shows that the key element appearing for an optimal…

统计理论 · 数学 2012-10-25 Matthieu Solnon , Sylvain Arlot , Francis Bach

We propose a penalized likelihood method to jointly estimate multiple precision matrices for use in quadratic discriminant analysis and model based clustering. A ridge penalty and a ridge fusion penalty are used to introduce shrinkage and…

机器学习 · 统计学 2014-05-06 Bradley S. Price , Charles J. Geyer , Adam J. Rothman

Structured prediction can be considered as a generalization of many standard supervised learning tasks, and is usually thought as a simultaneous prediction of multiple labels. One standard approach is to maximize a score function on the…

机器学习 · 计算机科学 2021-02-19 Kevin Bello , Asish Ghoshal , Jean Honorio

This article concerns the expressive power of depth in deep feed-forward neural nets with ReLU activations. Specifically, we answer the following question: for a fixed $d_{in}\geq 1,$ what is the minimal width $w$ so that neural nets with…

机器学习 · 统计学 2018-03-13 Boris Hanin , Mark Sellke

In adaptive data analysis, the user makes a sequence of queries on the data, where at each step the choice of query may depend on the results in previous steps. The releases are often randomized in order to reduce overfitting for such…

机器学习 · 统计学 2016-02-16 Yu-Xiang Wang , Jing Lei , Stephen E. Fienberg

We present a new active learning algorithm based on nonparametric estimators of the regression function. Our investigation provides probabilistic bounds for the rates of convergence of the generalization error achievable by proposed method…

统计理论 · 数学 2011-11-03 Stanislav Minsker

Quantile regression is the task of estimating a specified percentile response, such as the median, from a collection of known covariates. We study quantile regression with rectified linear unit (ReLU) neural networks as the chosen model…

统计理论 · 数学 2020-12-21 Oscar Hernan Madrid Padilla , Wesley Tansey , Yanzhen Chen

Recently, Forbes, Kumar and Saptharishi [CCC, 2016] proved that there exists an explicit $d^{O(1)}$-variate and degree $d$ polynomial $P_{d}\in VNP$ such that if any depth four circuit $C$ of bounded formal degree $d$ which computes a…

计算复杂性 · 计算机科学 2021-07-22 Suryajith Chillara

Kernel Stein discrepancies (KSDs) have emerged as a powerful tool for quantifying goodness-of-fit over the last decade, featuring numerous successful applications. To the best of our knowledge, all existing KSD estimators with known rate…

In the multivariate regression, also referred to as multi-task learning in machine learning, the goal is to recover a vector-valued function based on noisy observations. The vector-valued function is often assumed to be of low rank.…

统计理论 · 数学 2020-05-05 Wenjia Wang , Yi-Hui Zhou

We study minimax lower bounds for function estimation problems on large graph when the target function is smoothly varying over the graph. We derive minimax rates in the context of regression and classification problems on graphs that…

统计理论 · 数学 2018-02-16 Alisa Kirichenko , Harry van Zanten

We analyse adversarial bandit convex optimisation with an adversary that is restricted to playing functions of the form $f_t(x) = g_t(\langle x, \theta\rangle)$ for convex $g_t : \mathbb R \to \mathbb R$ and unknown $\theta \in \mathbb R^d$…

机器学习 · 计算机科学 2021-06-08 Tor Lattimore

In this work we consider the problem of estimating function-on-scalar regression models when the functions are observed over multi-dimensional or manifold domains and with potentially multivariate output. We establish the minimax rates of…

统计理论 · 数学 2019-02-21 Matthew Reimherr , Bharath Sriperumbudur , Hyun Bin Kang

This study extends power formulas proposed by Schochet (2008) assuming that the cluster-level score variable follows quadratic functional form. Results reveal that we need not be concerned with treatment by linear term interaction, and…

统计方法学 · 统计学 2020-06-09 Metin Bulus

From benign overfitting in overparameterized models to rich power-law scalings in performance, simple ridge regression displays surprising behaviors sometimes thought to be limited to deep neural networks. This balance of phenomenological…

机器学习 · 统计学 2026-05-08 Alexander Atanasov , Jacob A. Zavatone-Veth , Cengiz Pehlevan

Whether or not a local minimum of a cost function has a strongly convex neighborhood greatly influences the asymptotic convergence rate of optimizers. In this article, we rigorously analyze the prevalence of this property for the mean…

机器学习 · 计算机科学 2025-04-15 Felix Benning , Steffen Dereich