English
Related papers

Related papers: Ultrahigh Dimensional Feature Selection via Kernel…

200 papers

Given two sets of variables, derived from a common set of samples, sparse Canonical Correlation Analysis (CCA) seeks linear combinations of a small number of variables in each set, such that the induced canonical variables are maximally…

Machine Learning · Statistics 2016-05-31 Megasthenis Asteris , Anastasios Kyrillidis , Oluwasanmi Koyejo , Russell Poldrack

With the development of Earth observation technology, very-high-resolution (VHR) image has become an important data source of change detection. Nowadays, deep learning methods have achieved conspicuous performance in the change detection of…

Image and Video Processing · Electrical Eng. & Systems 2019-12-19 Chen Wu , Hongruixuan Chen , Bo Do , Liangpei Zhang

Canonical correlation analysis (CCA) is a multivariate statistical technique for finding the linear relationship between two sets of variables. The kernel generalization of CCA named kernel CCA has been proposed to find nonlinear relations…

Machine Learning · Statistics 2017-01-17 Xiaowei Zhang , Delin Chu , Li-Zhi Liao , Michael K. Ng

Random features approach has been widely used for kernel approximation in large-scale machine learning. A number of recent studies have explored data-dependent sampling of features, modifying the stochastic oracle from which random features…

Machine Learning · Computer Science 2021-11-03 Yinsong Wang , Shahin Shahrampour

Advancement in technology has generated abundant high-dimensional data that allows integration of multiple relevant studies. Due to their huge computational advantage, variable screening methods based on marginal correlation have become…

Methodology · Statistics 2017-10-12 Tianzhou Ma , Zhao Ren , George C. Tseng

Reducing the number of false discoveries is presently one of the most pressing issues in the life sciences. It is of especially great importance for many applications in neuroimaging and genomics, where datasets are typically…

Methodology · Statistics 2018-06-26 Alexej Gossmann , Pascal Zille , Vince Calhoun , Yu-Ping Wang

The applications of traditional statistical feature selection methods to high-dimension, low sample-size data often struggle and encounter challenging problems, such as overfitting, curse of dimensionality, computational infeasibility, and…

Machine Learning · Statistics 2023-12-19 Kexuan Li , Fangfang Wang , Lingli Yang , Ruiqi Liu

This paper proposes a novel model-free screening procedure for ultrahigh dimensional data analysis. By utilizing slicing technique which has been successfully ap- plied to continuous variables, we construct a new index called the fused…

Methodology · Statistics 2016-12-28 Yan Xiao-Dong , Xie Jin-Han , Ding Xian-Wen , Wang Zhi-Qiang , Tang Nian-Sheng

We study the problem of column selection in large-scale kernel canonical correlation analysis (KCCA) using the Nystr\"om approximation, where one approximates two positive semi-definite kernel matrices using "landmark" points from the…

Machine Learning · Computer Science 2016-02-09 Weiran Wang

Many relations of scientific interest are nonlinear, and even in linear systems distributions are often non-Gaussian, for example in fMRI BOLD data. A class of search procedures for causal relations in high dimensional data relies on sample…

Artificial Intelligence · Computer Science 2014-01-30 Joseph D. Ramsey

This paper presents a new efficient black-box attribution method based on Hilbert-Schmidt Independence Criterion (HSIC), a dependence measure based on Reproducing Kernel Hilbert Spaces (RKHS). HSIC measures the dependence between regions of…

Computer Vision and Pattern Recognition · Computer Science 2022-09-28 Paul Novello , Thomas Fel , David Vigouroux

This work proposes to learn fair low-rank tensor decompositions by regularizing the Canonical Polyadic Decomposition factorization with the kernel Hilbert-Schmidt independence criterion (KHSIC). It is shown, theoretically and empirically,…

Machine Learning · Computer Science 2021-04-29 Kevin Kim , Alex Gittens

Kernel and Multiple Kernel Canonical Correlation Analysis (CCA) are employed to classify schizophrenic and healthy patients based on their SNPs, DNA Methylation and fMRI data. Kernel and Multiple Kernel CCA are popular methods for finding…

Quantitative Methods · Quantitative Biology 2016-09-16 Owen Richfield , Md. Ashad Alam , Vince Calhoun , Yu-Ping Wang

Ultrahigh dimensional data sets are becoming increasingly prevalent in areas such as bioinformatics, medical imaging, and social network analysis. Sure independent screening of such data is commonly used to analyze such data. Nevertheless,…

Methodology · Statistics 2020-10-15 Randall Reese , Xiaotian Dai , Guifang Fu

As an unsupervised dimensionality reduction method, principal component analysis (PCA) has been widely considered as an efficient and effective preprocessing step for hyperspectral image (HSI) processing and analysis tasks. It takes each…

Computer Vision and Pattern Recognition · Computer Science 2018-09-18 Junjun Jiang , Jiayi Ma , Chen Chen , Zhongyuan Wang , Zhihua Cai , Lizhe Wang

This paper proposes a robust high-dimensional sparse canonical correlation analysis (CCA) method for investigating linear relationships between two high-dimensional random vectors, focusing on elliptical symmetric distributions. Traditional…

Methodology · Statistics 2025-04-18 Chengde Qian , Yanhong Liu , Long Feng

Kernel Principal Component Analysis (KPCA) is a key machine learning algorithm for extracting nonlinear features from data. In the presence of a large volume of high dimensional data collected in a distributed fashion, it becomes very…

Machine Learning · Computer Science 2016-02-16 Maria-Florina Balcan , Yingyu Liang , Le Song , David Woodruff , Bo Xie

Identifying significant subsets of the genes, gene shaving is an essential and challenging issue for biomedical research for a huge number of genes and the complex nature of biological networks,. Since positive definite kernel based methods…

Machine Learning · Statistics 2018-09-06 Md. Ashad Alam , Mohammad Shahjama , Md. Ferdush Rahman

Feature selection refers to the process of selecting useful features for machine learning tasks, and it is also a key step for structural health monitoring (SHM). This paper proposes a fast feature selection algorithm by efficiently…

Machine Learning · Statistics 2024-09-11 Sikai Zhang , Tingna Wang , Keith Worden , Limin Sun , Elizabeth J. Cross

Feature screening for ultrahigh-dimension, in general, proceeds with two essential steps. The first step is measuring and ranking the marginal dependence between response and covariates, and the second is determining the threshold. We…

Methodology · Statistics 2022-07-28 Linsui Deng , Yilin Zhang