English
Related papers

Related papers: Comparing the Pearson and Spearman Correlation Coe…

200 papers

Reparameterization (RP) and likelihood ratio (LR) gradient estimators are used throughout machine and reinforcement learning; however, they are usually explained as simple mathematical tricks without providing any insight into their nature.…

Machine Learning · Computer Science 2019-10-16 Paavo Parmas , Masashi Sugiyama

Mixing patterns in large self-organizing networks, such as the Internet, the World Wide Web, social and biological networks are often characterized by degree-degree dependencies between neighbouring nodes. In this paper we propose a new way…

Physics and Society · Physics 2015-06-04 Nelly Litvak , Remco van der Hofstad

Analysis of sample survey data often requires adjustments to account for missing data in the outcome variables of principal interest. Standard adjustment methods based on item imputation or on propensity weighting factors rely heavily on…

Methodology · Statistics 2016-03-08 Wei-Yin Loh , John Eltinge , MoonJung Cho , Yuanzhi Li

The Cox regression model is a popular model for analyzing the relationship between a covariate and a survival endpoint. The standard Cox model assumes a constant covariate effect across the entire covariate domain. However, in many…

Applications · Statistics 2019-09-02 Sarit Agami , David M. Zucker , Donna Spiegelman

Two typical forms of bias in user interaction data with recommender systems (RSs) are popularity bias and positivity bias, which manifest themselves as the over-representation of interactions with popular items or items that users prefer,…

Information Retrieval · Computer Science 2024-04-30 Jin Huang , Harrie Oosterhuis , Masoud Mansoury , Herke van Hoof , Maarten de Rijke

To date, it is still impossible to sample the entire mammalian brain with single-neuron precision. This forces one to either use spikes (focusing on few neurons) or to use coarse-sampled activity (averaging over many neurons, e.g. LFP).…

Neurons and Cognition · Quantitative Biology 2022-12-12 Joao Pinheiro Neto , Franz Paul Spitzner , Viola Priesemann

Insurance data can be asymmetric with heavy tails, causing inadequate adjustments of the usually applied models. To deal with this issue, hierarchical models for collective risk with heavy-tails of the claims distributions that take also…

Applications · Statistics 2021-01-26 Pamela M. Chiroque-Solano , Fernando A. S. Moura

The Peaks-Over Threshold is a fundamental method in the estimation of rare events such as small exceedance probabilities, extreme quantiles and return periods. The main problem with the Peaks-Over Threshold method relates to the selection…

Methodology · Statistics 2018-12-11 Richard Minkah , Tertius de Wet

Many evaluation measures are used to evaluate social biases in masked language models (MLMs). However, we find that these previously proposed evaluation measures are lacking robustness in scenarios with limited datasets. This is because…

Computation and Language · Computer Science 2024-01-23 Yang Liu

Random Field Theory has been used in the fMRI literature to address the multiple comparisons problem. The method provides an analytical solution for the computation of precise p-values when its assumptions are met. When its assumptions are…

Applications · Statistics 2016-07-28 Tim M. Tierney , Christopher A. Clark , David W. Carmichael

A margin-free measure of bivariate association generalizing Spearman's rho to the case of non-monotonic dependence is defined in terms of two square integrable functions on the unit interval. Properties of generalized Spearman correlation…

Methodology · Statistics 2025-12-12 Alexander J. McNeil , Johanna G. Neslehova , Andrew D. Smith

In observational studies, researchers must select a method to control for confounding. Options include propensity score methods and regression. It remains unclear how dataset characteristics (size, overlap in propensity scores, exposure…

Methodology · Statistics 2022-10-21 J. Wilkinson , M. A. Mamas , E. Kontopantelis

Recent advances have shown that statistical tests for the rank of cross-covariance matrices play an important role in causal discovery. These rank tests include partial correlation tests as special cases and provide further graphical…

Machine Learning · Computer Science 2025-06-13 Xinshuai Dong , Ignavier Ng , Boyang Sun , Haoyue Dai , Guang-Yuan Hao , Shunxing Fan , Peter Spirtes , Yumou Qiu , Kun Zhang

Covariate adjustment is desired by both practitioners and regulators of randomized clinical trials because it improves precision for estimating treatment effects. However, covariate adjustment presents a particular challenge in…

Methodology · Statistics 2023-07-20 Yunfan Li , Jessica L. Ross , Aaron M. Smith , David P. Miller

Financial networks based on Pearson correlations have been intensively studied. However, previous studies may have led to misleading and catastrophic results because of several critical shortcomings of the Pearson correlation. The local…

General Finance · Quantitative Finance 2025-12-04 Peng Liu

Propensity score (PS) methods are widely used in observational studies to reduce confounding and estimate causal treatment effects. However, the validity of PS-based causal estimators depends heavily on correct model specification, and…

Although conceptually related, variable selection and relative importance (RI) analysis have been treated quite differently in the literature. While RI is typically used for post-hoc model explanation, this paper explores its potential for…

Machine Learning · Statistics 2026-04-24 Tien-En Chang , Argon Chen

Sparse Representation (or coding) based Classification (SRC) has gained great success in face recognition in recent years. However, SRC emphasizes the sparsity too much and overlooks the correlation information which has been demonstrated…

Computer Vision and Pattern Recognition · Computer Science 2014-05-05 Jing Wang , Canyi Lu , Meng Wang , Peipei Li , Shuicheng Yan , Xuegang Hu

We study least squares linear regression over $N$ uncorrelated Gaussian features that are selected in order of decreasing variance. When the number of selected features $p$ is at most the sample size $n$, the estimator under consideration…

Statistics Theory · Mathematics 2019-10-04 Ji Xu , Daniel Hsu

Regressions are commonly used in environmental science and economics to identify causal or associative relationships between variables. In these settings, remote sensing-derived map products increasingly serve as sources of variables,…

Applications · Statistics 2025-07-04 Kerri Lu , Dan M. Kluger , Stephen Bates , Sherrie Wang