中文
相关论文

相关论文: Rank-based concordance for zero-inflated data: New…

200 篇论文

Low-rank matrix completion is the task of recovering unknown entries of a matrix by assuming that the true matrix admits a good low-rank approximation. Sometimes additional information about the variables is known, and incorporating this…

机器学习 · 计算机科学 2023-06-16 Benoît Loucheur , P. -A. Absil , Michel Journée

We study the rank of the instantaneous or spot covariance matrix $\Sigma_X(t)$ of a multidimensional continuous semi-martingale $X(t)$. Given high-frequency observations $X(i/n)$, $i=0,\ldots,n$, we test the null hypothesis…

统计理论 · 数学 2021-10-04 Markus Reiß , Lars Winkelmann

Low-Rank Adaptation (LoRA) has emerged as a widely adopted parameter-efficient fine-tuning (PEFT) technique for foundation models. Recent work has highlighted an inherent asymmetry in the initialization of LoRA's low-rank factors, which has…

Synchronization of rotations is the problem of estimating a set of rotations R_i in SO(n), i = 1, ..., N, based on noisy measurements of relative rotations R_i R_j^T. This fundamental problem has found many recent applications, most…

信息论 · 计算机科学 2016-01-07 Nicolas Boumal , Amit Singer , P. -A. Absil , Vincent D. Blondel

Integrating the summary statistics from genome-wide association study (\textsc{gwas}) and expression quantitative trait loci (e\textsc{qtl}) data provides a powerful way of identifying the genes whose expression levels are potentially…

统计方法学 · 统计学 2020-11-23 Rong Ma , T. Tony Cai , Hongzhe Li

Let $\mathbf{Q}=(Q_1,\ldots,Q_n)$ be a random vector drawn from the uniform distribution on the set of all $n!$ permutations of $\{1,2,\ldots,n\}$. Let $\mathbf{Z}=(Z_1,\ldots,Z_n)$, where $Z_j$ is the mean zero variance one random variable…

统计理论 · 数学 2015-11-18 Zhigang Bao , Liang-Ching Lin , Guangming Pan , Wang Zhou

Many data mining and statistical machine learning algorithms have been developed to select a subset of covariates to associate with a response variable. Spurious discoveries can easily arise in high-dimensional data analysis due to enormous…

统计理论 · 数学 2016-10-25 Jianqing Fan , Wen-Xin Zhou

In two influential contributions, Rosenbaum (2005, 2020) advocated for using the distances between component-wise ranks, instead of the original data values, to measure covariate similarity when constructing matching estimators of average…

统计理论 · 数学 2024-01-09 Matias D. Cattaneo , Fang Han , Zhexiao Lin

Estimating associations between spatial covariates and responses - rather than merely predicting responses - is central to environmental science, epidemiology, and economics. For instance, public health officials might be interested in…

机器学习 · 统计学 2025-11-11 David R. Burt , Renato Berlinghieri , Stephen Bates , Tamara Broderick

The Gini index is a popular inequality measure with many applications in social and economic studies. This paper studies semiparametric inference on the Gini indices of two semicontinuous populations. We characterize the distribution of…

统计理论 · 数学 2021-06-08 Meng Yuan , Pengfei Li , Changbao Wu

The analysis of Fermi Large Area Telescope (LAT) gamma-ray data in a given Region Of Interest (RoI) usually consists of performing a binned log-likelihood fit in order to determine the sky model that, after convolution with the instrument…

高能天体物理现象 · 物理学 2021-12-08 P. Bruel

Ideally, all analyses of normally distributed data should include the full covariance information between all data points. In practice, the full covariance matrix between all data points is not always available. Either because a result was…

统计方法学 · 统计学 2026-02-23 Lukas Koch

Sequencing-based technologies provide an abundance of high-dimensional biological datasets with skewed and zero-inflated measurements. Classification of such data with linear discriminant analysis leads to poor performance due to the…

统计方法学 · 统计学 2022-08-09 Hee Cheol Chung , Yang Ni , Irina Gaynanova

The reliability of machine learning systems critically assumes that the associations between features and labels remain similar between training and test distributions. However, unmeasured variables, such as confounders, break this…

机器学习 · 计算机科学 2020-08-17 Megha Srivastava , Tatsunori Hashimoto , Percy Liang

We consider the problem of statistical inference for ranking data, specifically rank aggregation, under the assumption that samples are incomplete in the sense of not comprising all choice alternatives. In contrast to most existing methods,…

机器学习 · 统计学 2017-12-05 Mohsen Ahmadi Fahandar , Eyke Hüllermeier , Inés Couso

While measures of concordance -- such as Spearman's rho, Kendall's tau, and Blomqvist's beta -- are continuous with respect to weak convergence, Chatterjee's rank correlation xi recently introduced in Azadkia and Chatterjee (2021) does not…

统计理论 · 数学 2026-04-14 Jonathan Ansari , Sebastian Fuchs

Rank-based dependence measures such as Spearman's footrule are robust and invariant, but they often fail to capture directional or asymmetric dependence in multivariate settings. This paper introduces a new family of directional Spearman's…

统计理论 · 数学 2026-01-27 Enrique de Amo , David García-Fernández , Manuel Úbeda-Flores

Spatial association measures for univariate static spatial data are widely used. When the data is in the form of a collection of spatial vectors with the same temporal domain of interest, we construct a measure of similarity between the…

统计方法学 · 统计学 2023-09-26 Divya Kappara , Arup Bose , Madhuchhanda Bhattacharjee

A key condition for obtaining reliable estimates of the causal effect of a treatment is overlap (a.k.a. positivity): the distributions of the features used to perform causal adjustment cannot be too different in the treated and control…

统计方法学 · 统计学 2021-04-14 Alexander D'Amour , Alexander Franks

The restricted polynomially-tilted pairwise interaction (RPPI) distribution gives a flexible model for compositional data. It is particularly well-suited to situations where some of the marginal distributions of the components of a…

统计方法学 · 统计学 2023-05-15 Janice L. Scealy , Kassel L. Hingee , John T. Kent , Andrew T. A. Wood