English
Related papers

Related papers: Quantifying the Fraction of Missing Information fo…

200 papers

We develop a new rank-based approach for univariate two-sample testing in the presence of missing data which makes no assumptions about the missingness mechanism. This approach is a theoretical extension of the Wilcoxon-Mann-Whitney test…

Methodology · Statistics 2024-03-25 Yijin Zeng , Niall M. Adams , Dean A. Bodenham

We derive a tight bound between the quality of estimating a quantum state by measurement and the success probability of undoing the measurement in arbitrary dimensional systems, which completely describes the tradeoff relation between the…

Quantum Physics · Physics 2012-10-12 Yong Wook Cheong , Seung-Woo Lee

This article develops an analytical framework for studying information divergences and likelihood ratios associated with Poisson processes and point patterns on general measurable spaces. The main results include explicit analytical…

Statistics Theory · Mathematics 2024-10-07 Lasse Leskelä

The correct use and interpretation of models depends on several steps, two of which being the calibration by parameter estimation and the analysis of uncertainty. In the biological literature, these steps are seldom discussed together, but…

Quantitative Methods · Quantitative Biology 2015-08-17 André Chalom , Paulo Inácio de Knegt López de Prado

High throughput metabolomics data are fraught with both non-ignorable missing observations and unobserved factors that influence a metabolite's measured concentration, and it is well known that ignoring either of these complications can…

Methodology · Statistics 2019-09-09 Chris McKennan , Carole Ober , Dan Nicolae

In this article, we investigate the asymptotic properties of Bayesian multiple testing procedures under general dependent setup, when the sample size and the number of hypotheses both tend to infinity. Specifically, we investigate strong…

Statistics Theory · Mathematics 2020-05-14 Noirrit Kiran Chandra , Sourabh Bhattacharya

Bayesian approaches have become increasingly popular in causal inference problems due to their conceptual simplicity, excellent performance and in-built uncertainty quantification ('posterior credible sets'). We investigate Bayesian…

Machine Learning · Statistics 2019-09-27 Kolyan Ray , Botond Szabo

Missing data are often dealt with multiple imputation. A crucial part of the multiple imputation process is selecting sensible models to generate plausible values for incomplete data. A method based on posterior predictive checking is…

Computation · Statistics 2026-05-14 Mingyang Cai , Stef van Buuren , Gerko Vink

Motivated by applications to goodness of fit testing, the empirical likelihood approach is generalized to allow for the number of constraints to grow with the sample size and for the constraints to use estimated criteria functions. The…

Statistics Theory · Mathematics 2013-07-24 Hanxiang Peng , Anton Schick

Existing sequential generalized estimating equation methodology for longitudinal and group-correlated data focuses on narrow hypotheses concerning treatment efficacy and often makes modeling assumptions that impede the desirable robustness…

Methodology · Statistics 2026-03-16 Nathan T. Provost , Abdus S. Wahed

In some multivariate problems with missing data, pairs of variables exist that are never observed together. For example, some modern biological tools can produce data of this form. As a result of this structure, the covariance matrix is…

Methodology · Statistics 2013-08-13 Max Grazier G'Sell , Shai S. Shen-Orr , Robert Tibshirani

Comparing two population means of network data is of paramount importance in a wide range of scientific applications. Many existing network inference solutions focus on global testing of entire networks, without comparing individual network…

Methodology · Statistics 2019-10-10 Yin Xia , Lexin Li

The main purpose of this paper is to introduce first a new family of empirical test statistics for testing a simple null hypothesis when the vector of parameters of interest are defined through a specific set of unbiased estimating…

Statistics Theory · Mathematics 2014-08-19 Narayanaswamy Balakrishnan , Nirian Martín , Leandro Pardo

We show that the Kullback-Leibler distance is a good measure of the statistical uncertainty of correlation matrices estimated by using a finite set of data. For correlation matrices of multivariate Gaussian variables we analytically…

Data Analysis, Statistics and Probability · Physics 2008-12-02 Michele Tumminello , Fabrizio Lillo , Rosario Nunzio Mantegna

Missing data is frequently encountered in many areas of statistics. Propensity score weighting is a popular method for handling missing data. The propensity score method employs a response propensity model, but correct specification of the…

Methodology · Statistics 2024-03-28 Hengfang Wang , Jae Kwang Kim , Jeongseop Han , Youngjo Lee

Many quantum chemical similarity measures have been derived and substantiated by applying concepts and quantities from information theory to the electron density. To justify the use of information theory, the electron density is usually…

Chemical Physics · Physics 2024-10-14 Guillaume Acke , Daria Van Hende , Ruben Van der Stichelen , Patrick Bultinck

Methods for addressing missing data have become much more accessible to applied researchers. However, little guidance exists to help researchers systematically identify plausible missing data mechanisms in order to ensure that these methods…

Applications · Statistics 2020-07-29 Adam Davey , Ting Dai

The analysis of clinical questionnaire data comes with many inherent challenges. These challenges include the handling of data with missing fields, as well as the overall interpretation of a dataset with many fields of different scales and…

Machine Learning · Computer Science 2021-08-04 Connor J. McLaughlin , Efi G. Kokkotou , Jean A. King , Lisa A. Conboy , Ali Yousefi

We live in unprecedented times in terms of our ability to use evidence to inform medical care. For example, we can perform data-driven post-test probability calculations. However, there is work to do. As has been previously noted,…

Other Statistics · Statistics 2025-07-29 Samuel J. Weisenthal , Amit K. Chowdhry

Incomplete observability of data generates an identification problem. There is no panacea for missing data. What one can learn about a population parameter depends on the assumptions one finds credible to maintain. The credibility of…

Econometrics · Economics 2022-05-17 Charles F. Manski