English
Related papers

Related papers: Sparse semiparametric canonical correlation analys…

200 papers

Gaussian factor models have proven widely useful for parsimoniously characterizing dependence in multivariate data. There is a rich literature on their extension to mixed categorical and continuous variables, using latent Gaussian variables…

Methodology · Statistics 2013-01-14 Jared S. Murray , David B. Dunson , Lawrence Carin , Joseph E. Lucas

Modern biomedical studies often collect multi-view data, that is, multiple types of data measured on the same set of objects. A popular model in high-dimensional multi-view data analysis is to decompose each view's data matrix into a…

Machine Learning · Statistics 2022-09-19 Hai Shu , Zhe Qu , Hongtu Zhu

We study the large sample properties of sparse M-estimators in the presence of pseudo-observations. Our framework covers a broad class of semi-parametric copula models, for which the marginal distributions are unknown and replaced by their…

Statistics Theory · Mathematics 2023-06-01 Jean-David Fermanian , Benjamin Poignard

Quantitative studies in many fields involve the analysis of multivariate data of diverse types, including measurements that we may consider binary, ordinal and continuous. One approach to the analysis of such mixed data is to use a copula…

Statistics Theory · Mathematics 2007-06-13 Peter D. Hoff

Recent developments in regularized Canonical Correlation Analysis (CCA) promise powerful methods for high-dimensional, multiview data analysis. However, justifying the structural assumptions behind many popular approaches remains a…

Methodology · Statistics 2025-11-18 Lennie Wells , Kumar Thurimella , Sergio Bacallado

Research on Poisson regression analysis for dependent data has been developed rapidly in the last decade. One of difficult problems in a multivariate case is how to construct a cross-correlation structure and at the meantime make sure that…

Methodology · Statistics 2017-10-05 A'yunin Sofro , Jian Qing Shi , Chunzheng Cao

Generalized canonical correlation analysis (GCCA) aims at finding latent low-dimensional common structure from multiple views (feature vectors in different domains) of the same entities. Unlike principal component analysis (PCA) that…

Machine Learning · Statistics 2017-08-02 Xiao Fu , Kejun Huang , Mingyi Hong , Nicholas D. Sidiropoulos , Anthony Man-Cho So

When facing multivariate covariates, general semiparametric regression techniques come at hand to propose flexible models that are unexposed to the curse of dimensionality. In this work a semiparametric copula-based estimator for…

Methodology · Statistics 2016-03-25 Mickael De Backer , Anouar El Ghouch , Ingrid Van Keilegom

Increasing epidemiologic evidence suggests that the diversity and composition of the gut microbiome can predict infection risk in cancer patients. Infections remain a major cause of morbidity and mortality during chemotherapy. Analyzing…

Methodology · Statistics 2025-05-20 Lei Wang , Yang Ni , Irina Gaynanova

We propose a new high dimensional semiparametric principal component analysis (PCA) method, named Copula Component Analysis (COCA). The semiparametric model assumes that, after unspecified marginally monotone transformations, the…

Machine Learning · Statistics 2014-02-20 Fang Han , Han Liu

An asymptotic behavior of canonical correlation analysis is studied when dimension d grows and the sample size n is fxed. In particular, we are interested in the conditions for which CCA works or fails in the HDLSS situation. This technical…

Statistics Theory · Mathematics 2016-09-14 Sungwon Lee

This paper studies the binary classification of two distributions with the same Gaussian copula in high dimensions. Under this semiparametric Gaussian copula setting, we derive an accurate semiparametric estimator of the log density ratio,…

Statistics Theory · Mathematics 2014-11-12 Yue Zhao , Marten Wegkamp

We consider the problem of testing for the presence of linear relationships between large sets of random variables based on a post-selection inference approach to canonical correlation analysis. The challenge is to adjust for the selection…

Methodology · Statistics 2020-10-20 Ian W. McKeague , Xin Zhang

Canonical Correlation Analysis (CCA) has been widely applied to jointly embed multiple views of data in a maximally correlated latent space. However, the alignment between various data perspectives, which is required by traditional…

Machine Learning · Computer Science 2023-12-11 Biqian Cheng , Evangelos E. Papalexakis , Jia Chen

Since the beginning of the 21st century, the size, breadth, and granularity of data in biology and medicine has grown rapidly. In the example of neuroscience, studies with thousands of subjects are becoming more common, which provide…

This paper studies high-dimensional canonical correlation analysis (CCA) with an emphasis on the vectors that define canonical variables. The paper shows that when two dimensions of data grow to infinity jointly and proportionally, the…

Econometrics · Economics 2025-01-24 Anna Bykhovskaya , Vadim Gorin

Integrative analyses of different high dimensional data types are becoming increasingly popular. Similarly, incorporating prior functional relationships among variables in data analysis has been a topic of increasing interest as it helps…

Methodology · Statistics 2016-06-09 Sandra E. Safo , Shuzhao Li , Qi Long

Modern large scale datasets are often plagued with missing entries. For tabular data with missing values, a flurry of imputation algorithms solve for a complete matrix which minimizes some penalized reconstruction error. However, almost…

Machine Learning · Statistics 2021-01-20 Yuxuan Zhao , Madeleine Udell

We propose novel first-order stochastic approximation algorithms for canonical correlation analysis (CCA). Algorithms presented are instances of inexact matrix stochastic gradient (MSG) and inexact matrix exponentiated gradient (MEG), and…

Machine Learning · Computer Science 2018-02-27 Raman Arora , Teodor V. Marinov , Poorya Mianjy , Nathan Srebro

Item nonresponse is frequently encountered in practice. Ignoring missing data can lose efficiency and lead to misleading inference. Fractional imputation is a frequentist approach of imputation for handling missing data. However, the…

Methodology · Statistics 2018-09-18 Hejian Sang , Jae Kwang Kim
‹ Prev 1 4 5 6 7 8 10 Next ›