English
Related papers

Related papers: Ultrahigh Dimensional Feature Selection via Kernel…

200 papers

We consider asymptotically exact inference on the leading canonical correlation directions and strengths between two high dimensional vectors under sparsity restrictions. In this regard, our main contribution is the development of a loss…

Statistics Theory · Mathematics 2022-02-10 Nilanjana Laha , Nathan Huey , Brent Coull , Rajarshi Mukherjee

Canonical Correlation Analysis (CCA) is a widespread technique for discovering linear relationships between two sets of variables $X \in \mathbb{R}^{n \times p}$ and $Y \in \mathbb{R}^{n \times q}$. In high dimensions however, standard…

Methodology · Statistics 2024-05-31 Claire Donnat , Elena Tuzhilina

Constraint-based causal discovery (CCD) algorithms require fast and accurate conditional independence (CI) testing. The Kernel Conditional Independence Test (KCIT) is currently one of the most popular CI tests in the non-parametric setting,…

Methodology · Statistics 2017-04-14 Eric V. Strobl , Kun Zhang , Shyam Visweswaran

Measuring and testing dependence between complex objects is of great importance in modern statistics. Most existing work relied on the distance between random variables, which inevitably required the moment conditions to guarantee the…

Methodology · Statistics 2023-04-19 Yilin Zhang , Songshan Yang

We propose the Sobolev Independence Criterion (SIC), an interpretable dependency measure between a high dimensional random variable X and a response variable Y . SIC decomposes to the sum of feature importance scores and hence can be used…

Machine Learning · Computer Science 2019-11-01 Youssef Mroueh , Tom Sercu , Mattia Rigotti , Inkit Padhi , Cicero Dos Santos

We propose a novel kernel based post selection inference (PSI) algorithm, which can not only handle non-linearity in data but also structured output such as multi-dimensional and multi-label outputs. Specifically, we develop a PSI algorithm…

Machine Learning · Statistics 2016-10-17 Makoto Yamada , Yuta Umezu , Kenji Fukumizu , Ichiro Takeuchi

We approach self-supervised learning of image representations from a statistical dependence perspective, proposing Self-Supervised Learning with the Hilbert-Schmidt Independence Criterion (SSL-HSIC). SSL-HSIC maximizes dependence between…

Machine Learning · Statistics 2021-12-06 Yazhe Li , Roman Pogodin , Danica J. Sutherland , Arthur Gretton

Recently the widely used multi-view learning model, Canonical Correlation Analysis (CCA) has been generalised to the non-linear setting via deep neural networks. Existing deep CCA models typically first decorrelate the feature dimensions of…

Computer Vision and Pattern Recognition · Computer Science 2018-03-28 Xiaobin Chang , Tao Xiang , Timothy M. Hospedales

Feature screening for ultra high dimensional feature spaces plays a critical role in the analysis of data sets whose predictors exponentially exceed the number of observations. Such data sets are becoming increasingly prevalent in areas…

Methodology · Statistics 2018-01-31 Randall Reese , Xiaotian Dai , Guifang Fu

We introduce the Kernel Calibration Conditional Stein Discrepancy test (KCCSD test), a non-parametric, kernel-based test for assessing the calibration of probabilistic models with well-defined scores. In contrast to previous methods, our…

Machine Learning · Statistics 2025-10-17 Pierre Glaser , David Widmann , Fredrik Lindsten , Arthur Gretton

We present Deep Tensor Canonical Correlation Analysis (DTCCA), a method to learn complex nonlinear transformations of multiple views (more than two) of data such that the resulting representations are linearly correlated in high order. The…

Machine Learning · Computer Science 2020-05-26 Hok Shing Wong , Li Wang , Raymond Chan , Tieyong Zeng

We propose the conditional predictive impact (CPI), a consistent and unbiased estimator of the association between one or several features and a given outcome, conditional on a reduced feature set. Building on the knockoff framework of…

Methodology · Statistics 2021-05-14 David S. Watson , Marvin N. Wright

This paper introduces the Class-wise Principal Component Analysis, a supervised feature extraction method for hyperspectral data. Hyperspectral Imaging (HSI) has appeared in various fields in recent years, including Remote Sensing.…

Computer Vision and Pattern Recognition · Computer Science 2021-04-12 Dimitra Koumoutsou , Eleni Charou , Georgios Siolas , Giorgos Stamou

Microarray studies, in order to identify genes associated with an outcome of interest, usually produce noisy measurements for a large number of gene expression features from a small number of subjects. One common approach to analyzing such…

Methodology · Statistics 2021-04-21 Linh Nghiem , Francis K. C. Hui , Samuel Mueller , A. H. Welsh

Canonical Correlation Analysis (CCA) is a multivariate technique that takes two datasets and forms the most highly correlated possible pairs of linear combinations between them. Each subsequent pair of linear combinations is orthogonal to…

Methodology · Statistics 2015-12-22 Jacob Coleman , Joseph Replogle , Gabriel Chandler , Johanna Hardin

Sparse Canonical Correlation Analysis (CCA) has received considerable attention in high-dimensional data analysis to study the relationship between two sets of random variables. However, there has been remarkably little theoretical…

Statistics Theory · Mathematics 2013-11-26 Mengjie Chen , Chao Gao , Zhao Ren , Harrison H. Zhou

Nonparametric feature selection in high-dimensional data is an important and challenging problem in statistics and machine learning fields. Most of the existing methods for feature selection focus on parametric or additive models which may…

Methodology · Statistics 2021-03-31 Hang Yu , Yuanjia Wang , Donglin Zeng

Independent Component Analysis (ICA) is a technique for unsupervised exploration of multi-channel data that is widely used in observational sciences. In its classic form, ICA relies on modeling the data as linear mixtures of non-Gaussian…

Machine Learning · Statistics 2018-08-01 Pierre Ablin , Jean-François Cardoso , Alexandre Gramfort

We introduce a new approach to variable selection, called Predictive Correlation Screening, for predictor design. Predictive Correlation Screening (PCS) implements false positive control on the selected variables, is well suited to small…

Machine Learning · Statistics 2013-04-11 Hamed Firouzi , Bala Rajaratnam , Alfred Hero

Ultra-high dimensional longitudinal data are increasingly common and the analysis is challenging both theoretically and methodologically. We offer a new automatic procedure for finding a sparse semivarying coefficient model, which is widely…

Methodology · Statistics 2014-09-24 Ming-Yen Cheng , Toshio Honda , Jialiang Li , Heng Peng
‹ Prev 1 3 4 5 6 7 10 Next ›