English
Related papers

Related papers: Rank-based concordance for zero-inflated data: New…

200 papers

Low-rank matrix completion is the task of recovering unknown entries of a matrix by assuming that the true matrix admits a good low-rank approximation. Sometimes additional information about the variables is known, and incorporating this…

Machine Learning · Computer Science 2023-06-16 Benoît Loucheur , P. -A. Absil , Michel Journée

We study the rank of the instantaneous or spot covariance matrix $\Sigma_X(t)$ of a multidimensional continuous semi-martingale $X(t)$. Given high-frequency observations $X(i/n)$, $i=0,\ldots,n$, we test the null hypothesis…

Statistics Theory · Mathematics 2021-10-04 Markus Reiß , Lars Winkelmann

Low-Rank Adaptation (LoRA) has emerged as a widely adopted parameter-efficient fine-tuning (PEFT) technique for foundation models. Recent work has highlighted an inherent asymmetry in the initialization of LoRA's low-rank factors, which has…

Machine Learning · Statistics 2025-06-18 Anastasis Kratsios , Tin Sum Cheng , Aurelien Lucchi , Haitz Sáez de Ocáriz Borde

Synchronization of rotations is the problem of estimating a set of rotations R_i in SO(n), i = 1, ..., N, based on noisy measurements of relative rotations R_i R_j^T. This fundamental problem has found many recent applications, most…

Information Theory · Computer Science 2016-01-07 Nicolas Boumal , Amit Singer , P. -A. Absil , Vincent D. Blondel

Integrating the summary statistics from genome-wide association study (\textsc{gwas}) and expression quantitative trait loci (e\textsc{qtl}) data provides a powerful way of identifying the genes whose expression levels are potentially…

Methodology · Statistics 2020-11-23 Rong Ma , T. Tony Cai , Hongzhe Li

Let $\mathbf{Q}=(Q_1,\ldots,Q_n)$ be a random vector drawn from the uniform distribution on the set of all $n!$ permutations of $\{1,2,\ldots,n\}$. Let $\mathbf{Z}=(Z_1,\ldots,Z_n)$, where $Z_j$ is the mean zero variance one random variable…

Statistics Theory · Mathematics 2015-11-18 Zhigang Bao , Liang-Ching Lin , Guangming Pan , Wang Zhou

Many data mining and statistical machine learning algorithms have been developed to select a subset of covariates to associate with a response variable. Spurious discoveries can easily arise in high-dimensional data analysis due to enormous…

Statistics Theory · Mathematics 2016-10-25 Jianqing Fan , Wen-Xin Zhou

In two influential contributions, Rosenbaum (2005, 2020) advocated for using the distances between component-wise ranks, instead of the original data values, to measure covariate similarity when constructing matching estimators of average…

Statistics Theory · Mathematics 2024-01-09 Matias D. Cattaneo , Fang Han , Zhexiao Lin

Estimating associations between spatial covariates and responses - rather than merely predicting responses - is central to environmental science, epidemiology, and economics. For instance, public health officials might be interested in…

Machine Learning · Statistics 2025-11-11 David R. Burt , Renato Berlinghieri , Stephen Bates , Tamara Broderick

The Gini index is a popular inequality measure with many applications in social and economic studies. This paper studies semiparametric inference on the Gini indices of two semicontinuous populations. We characterize the distribution of…

Statistics Theory · Mathematics 2021-06-08 Meng Yuan , Pengfei Li , Changbao Wu

The analysis of Fermi Large Area Telescope (LAT) gamma-ray data in a given Region Of Interest (RoI) usually consists of performing a binned log-likelihood fit in order to determine the sky model that, after convolution with the instrument…

High Energy Astrophysical Phenomena · Physics 2021-12-08 P. Bruel

Ideally, all analyses of normally distributed data should include the full covariance information between all data points. In practice, the full covariance matrix between all data points is not always available. Either because a result was…

Methodology · Statistics 2026-02-23 Lukas Koch

Sequencing-based technologies provide an abundance of high-dimensional biological datasets with skewed and zero-inflated measurements. Classification of such data with linear discriminant analysis leads to poor performance due to the…

Methodology · Statistics 2022-08-09 Hee Cheol Chung , Yang Ni , Irina Gaynanova

The reliability of machine learning systems critically assumes that the associations between features and labels remain similar between training and test distributions. However, unmeasured variables, such as confounders, break this…

Machine Learning · Computer Science 2020-08-17 Megha Srivastava , Tatsunori Hashimoto , Percy Liang

We consider the problem of statistical inference for ranking data, specifically rank aggregation, under the assumption that samples are incomplete in the sense of not comprising all choice alternatives. In contrast to most existing methods,…

Machine Learning · Statistics 2017-12-05 Mohsen Ahmadi Fahandar , Eyke Hüllermeier , Inés Couso

While measures of concordance -- such as Spearman's rho, Kendall's tau, and Blomqvist's beta -- are continuous with respect to weak convergence, Chatterjee's rank correlation xi recently introduced in Azadkia and Chatterjee (2021) does not…

Statistics Theory · Mathematics 2026-04-14 Jonathan Ansari , Sebastian Fuchs

Rank-based dependence measures such as Spearman's footrule are robust and invariant, but they often fail to capture directional or asymmetric dependence in multivariate settings. This paper introduces a new family of directional Spearman's…

Statistics Theory · Mathematics 2026-01-27 Enrique de Amo , David García-Fernández , Manuel Úbeda-Flores

Spatial association measures for univariate static spatial data are widely used. When the data is in the form of a collection of spatial vectors with the same temporal domain of interest, we construct a measure of similarity between the…

Methodology · Statistics 2023-09-26 Divya Kappara , Arup Bose , Madhuchhanda Bhattacharjee

A key condition for obtaining reliable estimates of the causal effect of a treatment is overlap (a.k.a. positivity): the distributions of the features used to perform causal adjustment cannot be too different in the treated and control…

Methodology · Statistics 2021-04-14 Alexander D'Amour , Alexander Franks

The restricted polynomially-tilted pairwise interaction (RPPI) distribution gives a flexible model for compositional data. It is particularly well-suited to situations where some of the marginal distributions of the components of a…

Methodology · Statistics 2023-05-15 Janice L. Scealy , Kassel L. Hingee , John T. Kent , Andrew T. A. Wood
‹ Prev 1 3 4 5 6 7 10 Next ›