English
Related papers

Related papers: Adjusting the adjusted Rand Index -- A multinomial…

200 papers

Quantifying out-of-sample discrimination performance for time-to-event outcomes is a fundamental step for model evaluation and selection in the context of predictive modelling. The concordance index, or C-index, is a widely used metric for…

Machine Learning · Statistics 2025-08-21 Begoña B. Sierra , Colin McLean , Peter S. Hall , Catalina A. Vallejos

To backpropagate the gradients through stochastic binary layers, we propose the augment-REINFORCE-merge (ARM) estimator that is unbiased, exhibits low variance, and has low computational complexity. Exploiting variable augmentation,…

Machine Learning · Statistics 2019-09-11 Mingzhang Yin , Mingyuan Zhou

We propose small-variance asymptotic approximations for the inference of tumor heterogeneity (TH) using next-generation sequencing data. Understanding TH is an important and open research problem in biology. The lack of appropriate…

Methodology · Statistics 2015-11-17 Yanxun Xu , Peter Mueller , Yuan Yuan , Kamalakar Gulukota , Yuan Ji

Traditionally, the Dirichlet-multinomial distribution has been recognized as a key model for contingency tables generated by cluster sampling schemes. There are, however, other possible distributions appropriate for these contingency…

Methodology · Statistics 2016-09-26 Juana M. Alonso-Revenga , Nirian Martin , Leandro Pardo

In this work we perform computational and analytical studies of the Randi\'c index $R(G)$ in Erd\"os-R\'{e}nyi models $G(n,p)$ characterized by $n$ vertices connected independently with probability $p \in (0,1)$. First, from a detailed…

Disordered Systems and Neural Networks · Physics 2020-02-11 C. T. Martinez-Martinez , J. A. Mendez-Bermudez , Jose M. Rodriguez , Jose M. Sigarreta

The Metropolis-Hastings algorithm allows one to sample asymptotically from any probability distribution $\pi$. There has been recently much work devoted to the development of variants of the MH update which can handle scenarios where such…

Computation · Statistics 2018-03-28 Christophe Andrieu , Arnaud Doucet , Sinan Yıldırım , Nicolas Chopin

Mutual information (MI) is a fundamental measure of statistical dependence, with a myriad of applications to information theory, statistics, and machine learning. While it possesses many desirable structural properties, the estimation of…

Information Theory · Computer Science 2021-10-19 Ziv Goldfeld , Kristjan Greenewald

Recent progress has been made with Adaptive Multiple Importance Sampling (AMIS) methods that show improvement in effective sample size. However, consistency for the AMIS estimator has only been established in very restricted cases.…

Optimization and Control · Mathematics 2018-03-22 Sep Thijssen , H. J. Kappen

A novel non-parametric estimator of the correlation between grouped measurements of a quantity is proposed in the presence of noise. This work is primarily motivated by functional brain network construction from fMRI data, where brain…

Methodology · Statistics 2023-02-16 Hanâ Lbath , Alexander Petersen , Wendy Meiring , Sophie Achard

In modern randomized experiments, large-scale data collection increasingly yields rich baseline covariates and auxiliary information from multiple sources. Such information offers opportunities for more precise treatment effect estimation,…

Methodology · Statistics 2026-03-10 Wei Ma , Zeqi Wu , Zheng Zhang

Complete randomization allows for consistent estimation of the average treatment effect based on the difference in means of the outcomes without strong modeling assumptions on the outcome-generating process. Appropriate use of the…

Methodology · Statistics 2021-08-03 Anqi Zhao , Peng Ding

Data augmentation plays a crucial role in enhancing the robustness and performance of machine learning models across various domains. In this study, we introduce a novel mixed-sample data augmentation method called RandoMix. RandoMix is…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Xiaoliang Liu , Furao Shen , Jian Zhao , Changhai Nie

We discuss how to handle matching-adjusted indirect comparison (MAIC) from a data analyst's perspective. We introduce several multivariate data analysis methods to assess the appropriateness of MAIC for a given data set. These methods focus…

Applications · Statistics 2022-03-18 Ekkehard Glimm , Lillian Yau

Variance reduction for causal inference in the presence of network interference is often achieved through either outcome modeling, typically analyzed under unit-randomized Bernoulli designs, or clustered experimental designs, typically…

Methodology · Statistics 2026-01-19 Matthew Eichhorn , Samir Khan , Johan Ugander , Christina Lee Yu

The multivariate normal density is a monotonic function of the distance to the mean, and its ellipsoidal shape is due to the underlying Euclidean metric. We suggest to replace this metric with a locally adaptive, smoothly changing…

Machine Learning · Statistics 2016-09-26 Georgios Arvanitidis , Lars Kai Hansen , Søren Hauberg

The contribution to pushing the boundaries of knowledge is a critical metric for evaluating the research performance of countries and institutions, which in many cases is not revealed by common bibliometric indicators. The Rk-index was…

Digital Libraries · Computer Science 2024-11-28 Alonso Rodriguez-Navarro

Early detection of anomalies in medical images such as brain MRI is highly relevant for diagnosis and treatment of many conditions. Supervised machine learning methods are limited to a small number of pathologies where there is good…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Alexander Frotscher , Jaivardhan Kapoor , Thomas Wolfers , Christian F. Baumgartner

There are multiple cluster randomised trial designs that vary in when the clusters cross between control and intervention states, when observations are made within clusters, and how many observations are made at that time point. Identifying…

Methodology · Statistics 2023-07-20 Samuel I. Watson , Alan Girling , Karla Hemming

Difference-in-differences is a widely-used evaluation strategy that draws causal inference from observational panel data. Its causal identification relies on the assumption of parallel trends, which is scale dependent and may be…

Applications · Statistics 2019-06-25 Peng Ding , Fan Li

The probabilistic Rand (PR) index has the following three problems: It lacks variations in its value over images; the normalized probabilistic Rand (NPR) index to address this is theoretically unclear, and the sampling method of pixel-pairs…

Computer Vision and Pattern Recognition · Computer Science 2019-06-20 Hisashi Shimodaira