English
Related papers

Related papers: Multivariate quantile-based permutation tests with…

200 papers

Weighting methods in causal inference have been widely used to achieve a desirable level of covariate balancing. However, the existing weighting methods have desirable theoretical properties only when a certain model, either the propensity…

Machine Learning · Statistics 2023-05-24 Insung Kong , Yuha Park , Joonhyuk Jung , Kwonsang Lee , Yongdai Kim

We consider statistical procedures for hypothesis testing of real valued functionals of matched pairs with missing values. In order to improve the accuracy of existing methods, we propose a novel multiplication combination procedure.…

Statistics Theory · Mathematics 2018-01-29 Lubna Amro , Frank Konietschke , Markus Pauly

A popular approach for testing if two univariate random variables are statistically independent consists of partitioning the sample space into bins, and evaluating a test statistic on the binned data. The partition size matters, and the…

Methodology · Statistics 2016-04-28 Ruth Heller , Yair Heller , Shachar Kaufman , Barak Brill , Malka Gorfine

Classical and more recent tests for detecting distributional changes in multivariate time series often lack power against alternatives that involve changes in the cross-sectional dependence structure. To be able to detect such changes…

Statistics Theory · Mathematics 2014-09-16 Axel Bücher , Ivan Kojadinovic , Tom Rohmer , Johan Segers

The topic of this paper is testing exchangeability using e-values in the batch mode, with the Markov model as alternative. The null hypothesis of exchangeability is formalized as a Kolmogorov-type compression model, and the Bayes mixture of…

Methodology · Statistics 2023-05-10 Vladimir Vovk

Coupling probability measures lies at the core of many problems in statistics and machine learning, from domain adaptation to transfer learning and causal inference. Yet, even when restricted to deterministic transports, such couplings are…

Machine Learning · Statistics 2025-09-22 Lucas De Lara , Luca Ganassali

For estimating the proportion of false null hypotheses in multiple testing, a family of estimators by Storey (2002) is widely used in the applied and statistical literature, with many methods suggested for selecting the parameter $\lambda$.…

Methodology · Statistics 2024-05-07 Anica Kostic , Piotr Fryzlewicz

We propose a new powerful family of tests of univariate normality. These tests are based on an initial value problem in the space of characteristic functions originating from the fixed point property of the normal distribution in the zero…

Statistics Theory · Mathematics 2020-02-28 Bruno Ebner

Motivation: Combining the results of different experiments to exhibit complex patterns or to improve statistical power is a typical aim of data integration. The starting point of the statistical analysis often comes as sets of p-values…

Methodology · Statistics 2021-12-02 Tristan Mary-Huard , Sarmistha Das , Indranil Mukhopadhyay , Stéphane Robin

In the big data era, the need to reevaluate traditional statistical methods is paramount due to the challenges posed by vast datasets. While larger samples theoretically enhance accuracy and hypothesis testing power without increasing false…

Methodology · Statistics 2026-01-09 Xuekui Zhang , Li Xing , Jing Zhang , Soojeong Kim

A common problem in genetics is that of testing whether a set of highly dependent gene expressions differ between two populations, typically in a high-dimensional setting where the data dimension is larger than the sample size. Most…

Methodology · Statistics 2015-03-11 Måns Thulin

Many important problems in psychology and biomedical studies require testing for overdispersion, correlation and heterogeneity in mixed effects and latent variable models, and score tests are particularly useful for this purpose. But the…

Statistics Theory · Mathematics 2007-06-13 Hongtu Zhu , Heping Zhang

Stochastic dominance is an important concept in probability theory, econometrics and social choice theory for robustly modeling agents' preferences between random outcomes. While many works have been dedicated to the univariate case, little…

Machine Learning · Statistics 2024-06-11 Gabriel Rioux , Apoorva Nitsure , Mattia Rigotti , Kristjan Greenewald , Youssef Mroueh

Optimal transport and its related problems, including optimal partial transport, have proven to be valuable tools in machine learning for computing meaningful distances between probability or positive measures. This success has led to a…

Machine Learning · Computer Science 2023-07-26 Xinran Liu , Yikun Bai , Huy Tran , Zhanqi Zhu , Matthew Thorpe , Soheil Kolouri

Recently, optimal transport-based approaches have gained attention for deriving counterfactuals, e.g., to quantify algorithmic discrimination. However, in the general multivariate setting, these methods are often opaque and difficult to…

Machine Learning · Computer Science 2025-05-22 Agathe Fernandes Machado , Arthur Charpentier , Ewen Gallic

Consider a big data multiple testing task, where, due to storage and computational bottlenecks, one is given a very large collection of p-values by splitting into manageable chunks and distributing over thousands of computer nodes. This…

Methodology · Statistics 2018-05-08 Subhadeep Mukhopadhyay

Matching on covariates is a well-established framework for estimating causal effects in observational studies. The principal challenge stems from the often high-dimensional structure of the problem. Many methods have been introduced to…

Methodology · Statistics 2022-07-12 Florian Gunsilius , Yuliang Xu

For general repeated measures designs the Wald-type statistic (WTS) is an asymptotically valid procedure allowing for unequal covariance matrices and possibly non-normal multivariate observations. The drawback of this procedure is the poor…

Methodology · Statistics 2016-06-24 Sarah Friedrich , Edgar Brunner , Markus Pauly

The univariate quantile-quantile (Q-Q) plot is a well-known graphical tool for examining whether two data sets are generated from the same distribution or not. It is also used to determine how well a specified probability distribution fits…

Statistics Theory · Mathematics 2014-07-07 Subhra Sankar Dhar , Biman Chakraborty , Probal Chaudhuri

We consider goodness-of-fit tests with i.i.d. samples generated from a categorical distribution $(p_1,...,p_k)$. For a given $(q_1,...,q_k)$, we test the null hypothesis whether $p_j=q_{\pi(j)}$ for some label permutation $\pi$. The…

Statistics Theory · Mathematics 2018-07-30 Chao Gao