English
Related papers

Related papers: On Single Index Models beyond Gaussian Data

200 papers

We present an algorithm to identify sparse dependence structure in continuous and non-Gaussian probability distributions, given a corresponding set of data. The conditional independence structure of an arbitrary distribution can be…

Machine Learning · Computer Science 2017-11-07 Rebecca E. Morrison , Ricardo Baptista , Youssef Marzouk

Gaussian processes are a powerful framework for quantifying uncertainty and for sequential decision-making but are limited by the requirement of solving linear systems. In general, this has a cubic cost in dataset size and is sensitive to…

We provide efficient algorithms for the problem of distribution learning from high-dimensional Gaussian data where in each sample, some of the variable values are missing. We suppose that the variables are missing not at random (MNAR). The…

Machine Learning · Computer Science 2025-04-29 Arnab Bhattacharyya , Constantinos Daskalakis , Themis Gouleakis , Yuhao Wang

We focus on the task of learning a single index model $\sigma(w^\star \cdot x)$ with respect to the isotropic Gaussian distribution in $d$ dimensions. Prior work has shown that the sample complexity of learning $w^\star$ is governed by the…

Machine Learning · Computer Science 2023-05-19 Alex Damian , Eshaan Nichani , Rong Ge , Jason D. Lee

In this work we consider generic Gaussian Multi-index models, in which the labels only depend on the (Gaussian) $d$-dimensional inputs through their projection onto a low-dimensional $r = O_d(1)$ subspace, and we study efficient agnostic…

Machine Learning · Computer Science 2025-06-09 Alex Damian , Jason D. Lee , Joan Bruna

We consider the basic problem of learning Single-Index Models with respect to the square loss under the Gaussian distribution in the presence of adversarial label noise. Our main contribution is the first computationally efficient algorithm…

Machine Learning · Computer Science 2025-08-07 Puqian Wang , Nikos Zarifis , Ilias Diakonikolas , Jelena Diakonikolas

This paper introduces a new sparse spatio-temporal structured Gaussian process regression framework for online and offline Bayesian inference. This is the first framework that gives a time-evolving representation of the interdependencies…

Machine Learning · Statistics 2018-08-01 Danil Kuzin , Olga Isupova , Lyudmila Mihaylova

Stochastic gradient descent (SGD) is a cornerstone algorithm for high-dimensional optimization, renowned for its empirical successes. Recent theoretical advances have provided a deep understanding of how SGD enables feature learning in…

Machine Learning · Statistics 2026-02-23 Nived Rajaraman , Yanjun Han

Let $(\bX, Y)$ be a random pair taking values in $\mathbb R^p \times \mathbb R$. In the so-called single-index model, one has $Y=f^{\star}(\theta^{\star T}\bX)+\bW$, where $f^{\star}$ is an unknown univariate measurable function,…

Statistics Theory · Mathematics 2013-06-18 Pierre Alquier , Gérard Biau

Inference in Gaussian process (GP) models is computationally challenging for large data, and often difficult to approximate with a small number of inducing points. We explore an alternative approximation that employs stochastic inference…

Machine Learning · Statistics 2019-05-28 Jiaxin Shi , Mohammad Emtiyaz Khan , Jun Zhu

Extracting meaningful information from high-dimensional data poses a formidable modeling challenge, particularly when the data is obscured by noise or represented through different modalities. This research proposes a novel non-parametric…

Machine Learning · Computer Science 2024-08-27 Navid Ziaei , Behzad Nazari , Uri T. Eden , Alik Widge , Ali Yousefi

Given a single observation from a Gaussian distribution with unknown mean $\theta$, we design computationally efficient procedures that can approximately generate an observation from a different target distribution $Q_{\theta}$ uniformly…

Statistics Theory · Mathematics 2025-10-09 Mengqi Lou , Guy Bresler , Ashwin Pananjady

The information exponent ([BAGJ21]) and its extensions -- which are equivalent to the lowest degree in the Hermite expansion of the link function (after a potential label transform) for Gaussian single-index models -- have played an…

Machine Learning · Computer Science 2025-10-07 Yunwei Ren , Jason D. Lee

Unsupervised learning aims at the discovery of hidden structure that drives the observations in the real world. It is essential for success in modern machine learning. Latent variable models are versatile in unsupervised learning and have…

Machine Learning · Computer Science 2016-06-13 Furong Huang

I propose a novel framework that integrates stochastic differential equations (SDEs) with deep generative models to improve uncertainty quantification in machine learning applications involving structured and temporal data. This approach,…

Machine Learning · Statistics 2026-01-09 James Rice

We introduce a novel procedure that, given sparse data generated from a stationary deterministic nonlinear dynamical system, can characterize specific local and/or global dynamic behavior with rigorous probability guarantees. More…

Dynamical Systems · Mathematics 2023-09-19 Bogdan Batko , Marcio Gameiro , Ying Hung , William Kalies , Konstantin Mischaikow , Ewerton Vieira

This paper studies an asymptotic framework for conducting inference on parameters of the form $\phi(\theta_0)$, where $\phi$ is a known directionally differentiable function and $\theta_0$ is estimated by $\hat \theta_n$. In these settings,…

Statistics Theory · Mathematics 2016-01-14 Zheng Fang , Andres Santos

Graphical models are commonly used to represent conditional dependence relationships between variables. There are multiple methods available for exploring them from high-dimensional data, but almost all of them rely on the assumption that…

Machine Learning · Statistics 2020-04-22 Tianxi Li , Cheng Qian , Elizaveta Levina , Ji Zhu

A variational inference-based framework for training a multi-output Gaussian process latent variable model, specifically tailored to the tails-up spatio-temporal stream network, is developed. Training, given a censored observational data…

Methodology · Statistics 2026-05-21 Marno Basson , Tobias M. Louw , Theresa R. Smith

Even pruned by the state-of-the-art network compression methods, Graph Neural Networks (GNNs) training upon non-Euclidean graph data often encounters relatively higher time costs, due to its irregular and nasty density properties, compared…

Machine Learning · Computer Science 2022-10-04 Chunhui Zhang , Chao Huang , Yijun Tian , Qianlong Wen , Zhongyu Ouyang , Youhuan Li , Yanfang Ye , Chuxu Zhang