中文
相关论文

相关论文: Learning single-index models via harmonic decompos…

200 篇论文

We study the problem of learning multi-index models (MIMs), where the label depends on the input $\boldsymbol{x} \in \mathbb{R}^d$ only through an unknown $\mathsf{s}$-dimensional projection $\boldsymbol{W}_*^\mathsf{T} \boldsymbol{x} \in…

统计理论 · 数学 2026-02-11 Hugo Latourelle-Vigeant , Theodor Misiakiewicz

We focus on the task of learning a single index model $\sigma(w^\star \cdot x)$ with respect to the isotropic Gaussian distribution in $d$ dimensions. Prior work has shown that the sample complexity of learning $w^\star$ is governed by the…

机器学习 · 计算机科学 2023-05-19 Alex Damian , Eshaan Nichani , Rong Ge , Jason D. Lee

A single-index model (SIM) is a function of the form $\sigma(\mathbf{w}^{\ast} \cdot \mathbf{x})$, where $\sigma: \mathbb{R} \to \mathbb{R}$ is a known link function and $\mathbf{w}^{\ast}$ is a hidden unit vector. We study the task of…

机器学习 · 计算机科学 2024-11-11 Puqian Wang , Nikos Zarifis , Ilias Diakonikolas , Jelena Diakonikolas

The problem of statistical inference for regression coefficients in a high-dimensional single-index model is considered. Under elliptical symmetry, the single index model can be reformulated as a proxy linear model whose regression…

统计理论 · 数学 2021-03-02 Hamid Eftekhari , Moulinath Banerjee , Ya'acov Ritov

Sparse high-dimensional functions have arisen as a rich framework to study the behavior of gradient-descent methods using shallow neural networks, showcasing their ability to perform feature learning beyond linear models. Amongst those…

机器学习 · 计算机科学 2023-10-26 Joan Bruna , Loucas Pillaud-Vivien , Aaron Zweig

Hamiltonian Learning is a process of recovering system Hamiltonian from measurements, which is a fundamental problem in quantum information processing. In this study, we investigate the problem of learning the symmetric Hamiltonian from its…

量子物理 · 物理学 2025-11-04 Jing Zhou , D. L. Zhou

We study the problem of gradient descent learning of a single-index target function $f_*(\boldsymbol{x}) = \textstyle\sigma_*\left(\langle\boldsymbol{x},\boldsymbol{\theta}\rangle\right)$ under isotropic Gaussian data in $\mathbb{R}^d$,…

机器学习 · 计算机科学 2024-12-24 Jason D. Lee , Kazusato Oko , Taiji Suzuki , Denny Wu

Few neural architectures lend themselves to provable learning with gradient based methods. One popular model is the single-index model, in which labels are produced by composing an unknown linear projection with a possibly unknown scalar…

机器学习 · 计算机科学 2023-10-04 Aaron Zweig , Joan Bruna

We study the fundamental problem of learning a single neuron, i.e., a function of the form $\mathbf{x}\mapsto\sigma(\mathbf{w}\cdot\mathbf{x})$ for monotone activations $\sigma:\mathbb{R}\mapsto\mathbb{R}$, with respect to the $L_2^2$-loss…

机器学习 · 计算机科学 2022-06-20 Ilias Diakonikolas , Vasilis Kontonis , Christos Tzamos , Nikos Zarifis

Reflectional symmetry is ubiquitous in nature. While extrinsic reflectional symmetry can be easily parametrized and detected, intrinsic symmetry is much harder due to the high solution space. Previous works usually solve this problem by…

图形学 · 计算机科学 2019-11-04 Yi-Ling Qiao , Lin Gao , Shu-Zhi Liu , Ligang Liu , Yu-Kun Lai , Xilin Chen

The problem of recovering a structured signal $\mathbf{x} \in \mathbb{C}^p$ from a set of dimensionality-reduced linear measurements $\mathbf{b} = \mathbf {A}\mathbf {x}$ arises in a variety of applications, such as medical imaging,…

信息论 · 计算机科学 2016-05-25 Luca Baldassarre , Yen-Huan Li , Jonathan Scarlett , Baran Gözcü , Ilija Bogunovic , Volkan Cevher

Regression with a spherical response is challenging due to the absence of linear structure, making standard regression models inadequate. Existing methods, mainly parametric, lack the flexibility to capture the complex relationship induced…

统计方法学 · 统计学 2025-04-01 Houren Hong , Janice L. Scealy , Andrew T. A. Wood , Yanrong Yang

The information exponent ([BAGJ21]) and its extensions -- which are equivalent to the lowest degree in the Hermite expansion of the link function (after a potential label transform) for Gaussian single-index models -- have played an…

机器学习 · 计算机科学 2025-10-07 Yunwei Ren , Jason D. Lee

We consider the basic problem of learning Single-Index Models with respect to the square loss under the Gaussian distribution in the presence of adversarial label noise. Our main contribution is the first computationally efficient algorithm…

机器学习 · 计算机科学 2025-08-07 Puqian Wang , Nikos Zarifis , Ilias Diakonikolas , Jelena Diakonikolas

Single Index Models (SIMs) are simple yet flexible semi-parametric models for machine learning, where the response variable is modeled as a monotonic function of a linear combination of features. Estimation in this context requires learning…

机器学习 · 统计学 2016-12-01 Nikhil Rao , Ravi Ganti , Laura Balzano , Rebecca Willett , Robert Nowak

Homomorphic sensing is a recent algebraic-geometric framework that studies the unique recovery of points in a linear subspace from their images under a given collection of linear maps. It has been successful in interpreting such a recovery…

机器学习 · 计算机科学 2022-09-20 Liangzu Peng , Manolis C. Tsakiris

Single-index models are a class of functions given by an unknown univariate ``link'' function applied to an unknown one-dimensional projection of the input. These models are particularly relevant in high dimension, when the data might…

机器学习 · 计算机科学 2022-10-28 Alberto Bietti , Joan Bruna , Clayton Sanford , Min Jae Song

A recent line of research termed unlabeled sensing and shuffled linear regression has been exploring under great generality the recovery of signals from subsampled and permuted measurements; a challenging problem in diverse fields of data…

信息论 · 计算机科学 2019-07-19 Manolis C. Tsakiris , Liangzu Peng

Stochastic gradient descent (SGD) is a cornerstone algorithm for high-dimensional optimization, renowned for its empirical successes. Recent theoretical advances have provided a deep understanding of how SGD enables feature learning in…

机器学习 · 统计学 2026-02-23 Nived Rajaraman , Yanjun Han

Let $(\bX, Y)$ be a random pair taking values in $\mathbb R^p \times \mathbb R$. In the so-called single-index model, one has $Y=f^{\star}(\theta^{\star T}\bX)+\bW$, where $f^{\star}$ is an unknown univariate measurable function,…

统计理论 · 数学 2013-06-18 Pierre Alquier , Gérard Biau
‹ 上一页 1 2 3 10 下一页 ›