English
Related papers

Related papers: Affinity-based measures of medical diagnostic test…

200 papers

Matching is a commonly used causal inference study design in observational studies. Through matching on measured confounders between different treatment groups, valid randomization inferences can be conducted under the no unmeasured…

Methodology · Statistics 2024-09-20 Jeffrey Zhang , Siyu Heng

The problem of finding an appropriate geometrical/physical index for measuring a degree of inhomogeneity for a given space-time manifold is posed. Interrelations with the problem of understanding the gravitational/informational entropy are…

General Relativity and Quantum Cosmology · Physics 2007-05-23 Roustam Zalaletdinov

In this paper, we revisit the classical goodness-of-fit problems for univariate distributions; we propose a new testing procedure based on a characterisation of the uniform distribution. Asymptotic theory for the simple hypothesis case is…

Methodology · Statistics 2021-08-17 Bruno Ebner , Shawn Liebenberg , Jaco Visagie

In this paper, we tackle the problem of measuring similarity among graphs that represent real objects with noisy data. To account for noise, we relax the definition of similarity using the maximum weighted co-$k$-plex relaxation method,…

Data Structures and Algorithms · Computer Science 2016-01-26 Maritza Hernandez , Arman Zaribafiyan , Maliheh Aramon , Mohammad Naghibi

This paper proposes a new graph proximity measure. This measure is a derivative of network reliability. By analyzing its properties and comparing it against other proximity measures through graph examples, we demonstrate that it is more…

Social and Information Networks · Computer Science 2016-12-23 Haifeng Qian , Hui Wan , Mark N. Wegman , Luis A. Lastras , Ruchir Puri

When assessing the causal effect of a binary exposure using observational data, confounder imbalance across exposure arms must be addressed. Matching methods, including propensity score-based matching, can be used to deconfound the causal…

Methodology · Statistics 2024-10-01 Ernesto Ulloa-Pérez , Marco Carone , Alex Luedtke

In regression analysis, associations between continuous predictors and the outcome are often assumed to be linear. However, modeling the associations as non-linear can improve model fit. Many flexible modeling techniques, like (fractional)…

Repeated measurements are common in many fields, where random variables are observed repeatedly across different subjects. Such data have an underlying hierarchical structure, and it is of interest to learn covariance/correlation at…

Methodology · Statistics 2023-06-13 Sunpeng Duan , Guo Yu , Juntao Duan , Yuedong Wang

Measurement involves the determination of quantitative estimates of physical quantities from experiment, along with estimates of their associated uncertainties. Herewith an experimental system model is the key to extracting information from…

Applications · Statistics 2008-09-01 Vladimir B. Bokov

We study ratio metrics in A/B testing at the presence of correlation among observations coming from the same user and provides practical guidance especially when two metrics contradict each other. We propose new estimating methods to…

Applications · Statistics 2020-07-24 Keyu Nie , Yinfei Kong , Ted Tao Yuan , Pauline Berry Burke

Measures of similarity (or dissimilarity) are a key ingredient to many machine learning algorithms. We introduce DID, a pairwise dissimilarity measure applicable to a wide range of data spaces, which leverages the data's internal structure…

Machine Learning · Statistics 2022-03-08 Théophile Cantelobre , Carlo Ciliberto , Benjamin Guedj , Alessandro Rudi

The goals in clinical and cohort studies often include evaluation of the association of a time-dependent binary treatment or exposure with a survival outcome. Recently, several impactful studies targeted the association between…

Applications · Statistics 2019-01-24 Daniel Nevo , Tsuyoshi Hamada , Shuji Ogino , Molin Wang

We propose a semantic similarity metric for image registration. Existing metrics like Euclidean Distance or Normalized Cross-Correlation focus on aligning intensity values, giving difficulties with low intensity contrast or noise. Our…

Machine Learning · Computer Science 2021-04-21 Steffen Czolbe , Oswin Krause , Aasa Feragen

Matching on a low dimensional vector of scalar covariates consists of constructing groups of individuals in which each individual in a group is within a pre-specified distance from an individual in another group. However, matching in high…

Applications · Statistics 2024-05-06 Marina Hernandez , Ciprian Crainiceanu

We introduce a new phenomenological tool based on momentum region indicators to guide the analysis and interpretation of semi-inclusive deep-inelastic scattering measurements. The new tool, referred to as "affinity", is devised to help…

High Energy Physics - Phenomenology · Physics 2022-01-31 M. Boglione , M. Diefenthaler , S. Dolan , L. Gamberg , W. Melnitchouk , D. Pitonyak , A. Prokudin , N. Sato , Z. Scalyer

We propose the use of non-parametric, graph-based tests to assess the distributional balance of covariates in observational studies with multi-valued treatments. Our tests utilize graph structures ranging from Hamiltonian paths that connect…

Methodology · Statistics 2022-08-11 Eric A. Dunipace

Multivariate meta-analysis of test accuracy studies when tests are evaluated in terms of sensitivity and specificity at more than one threshold represents an effective way to synthesize results by fully exploiting the data, if compared to…

Methodology · Statistics 2019-01-29 Annamaria Guolo , Duc Khanh To

A ubiquitous feature of data of our era is their extra-large sizes and dimensions. Analyzing such high-dimensional data poses significant challenges, since the feature dimension is often much larger than the sample size. This thesis…

Statistics Theory · Mathematics 2025-09-11 Kai Yang

Computer models are commonly used to represent a wide range of real systems, but they often involve some unknown parameters. Estimating the parameters by collecting physical data becomes essential in many scientific fields, ranging from…

Applications · Statistics 2020-05-27 Chih-Li Sung , Beau David Barber , Berkley J. Walker

Several performance measures can be used for evaluating classification results: accuracy, F-measure, and many others. Can we say that some of them are better than others, or, ideally, choose one measure that is best in all situations? To…

Machine Learning · Computer Science 2022-01-25 Martijn Gösgens , Anton Zhiyanov , Alexey Tikhonov , Liudmila Prokhorenkova