English
Related papers

Related papers: Identifying Graphical Models

200 papers

We propose a method for detecting differential gene expression that exploits the correlation between genes. Our proposal averages the univariate scores of each feature with the scores in correlation neighborhoods. In a number of real and…

Statistics Theory · Mathematics 2007-06-13 Robert Tibshirani , Larry Wasserman

We address the problem of detecting changes in multivariate datastreams, and we investigate the intrinsic difficulty that change-detection methods have to face when the data dimension scales. In particular, we consider a general approach…

Machine Learning · Statistics 2017-12-04 Cesare Alippi , Giacomo Boracchi , Diego Carrera , Manuel Roveri

A learned generative model often produces biased statistics relative to the underlying data distribution. A standard technique to correct this bias is importance sampling, where samples from the model are weighted by the likelihood ratio…

Machine Learning · Statistics 2019-11-05 Aditya Grover , Jiaming Song , Alekh Agarwal , Kenneth Tran , Ashish Kapoor , Eric Horvitz , Stefano Ermon

AIC is commonly used for model selection but the precise value of AIC has no direct interpretation. We are interested in quantifying a difference of risks between two models. This may be useful for both an explanatory point of view or for…

Methodology · Statistics 2008-07-28 D. Commenges , A. Sayyareh , L. Letenneur , J. Guedj , A. Bar-Hen

Recent results in coupled or temporal graphical models offer schemes for estimating the relationship structure between features when the data come from related (but distinct) longitudinal sources. A novel application of these ideas is for…

Machine Learning · Statistics 2017-11-22 Ronak Mehta , Hyunwoo J. Kim , Shulei Wang , Sterling C. Johnson , Ming Yuan , Vikas Singh

Wide conditions are provided to guarantee asymptotic unbiasedness and L^2-consistency of the introduced estimates of the Kullback-Leibler divergence for probability measures in R^d having densities w.r.t. the Lebesgue measure. These…

Statistics Theory · Mathematics 2019-07-02 Alexander Bulinski , Denis Dimitrov

In this paper, the defining properties of a valid measure of the dependence between two random variables are reviewed and complemented with two original ones, shown to be more fundamental than other usual postulates. While other popular…

Methodology · Statistics 2019-12-03 Gery Geenens , Pierre Lafaye de Micheaux

Efficient textual data distributions (TDD) alignment and generation are open research problems in textual analytics and NLP. It is presently difficult to parsimoniously and methodologically confirm that two or more natural language datasets…

Computation and Language · Computer Science 2021-07-06 Jim Samuel , Ratnakar Palle , Eduardo Correa Soares

Group sequential designs enable interim analyses and potential early stopping for efficacy or futility. While these adaptations improve trial efficiency and ethical considerations, they also introduce bias into the adapted analyses. We…

Methodology · Statistics 2025-10-07 G. Caruso , W. F. Rosenberger , P. Mozgunov , N. Flournoy

Estimating the Shannon entropy of a discrete distribution from which we have only observed a small sample is challenging. Estimating other information-theoretic metrics, such as the Kullback-Leibler divergence between two sparsely sampled…

Data Analysis, Statistics and Probability · Physics 2023-02-24 Angelo Piga , Lluc Font-Pomarol , Marta Sales-Pardo , Roger Guimerà

Distance correlation is a measure of dependence between two paired random vectors or matrices of arbitrary, not necessarily equal, dimensions. Unlike Pearson correlation, the population distance correlation coefficient is zero if and only…

Methodology · Statistics 2025-06-19 Kontemeniotis Nikolaos , Vargiakakis Rafail , Tsagris Michail

Selecting an appropriate divergence measure is a critical aspect of machine learning, as it directly impacts model performance. Among the most widely used, we find the Kullback-Leibler (KL) divergence, originally introduced in kinetic…

Mathematical Physics · Physics 2025-07-16 Gennaro Auricchio , Giovanni Brigati , Paolo Giudici , Giuseppe Toscani

Proper scoring rules evaluate the quality of probabilistic predictions, playing an essential role in the pursuit of accurate and well-calibrated models. Every proper score decomposes into two fundamental components -- proper calibration…

Machine Learning · Computer Science 2023-12-15 Teodora Popordanoska , Sebastian G. Gruber , Aleksei Tiulpin , Florian Buettner , Matthew B. Blaschko

Both classical and respectively quantum observables can be modeled as somewhat similar examples of random variables. In such a model the associated measurements preserve the values spectrum of an observable but change the corresponding…

Statistical Mechanics · Physics 2008-03-20 S. Dumitru , A. Boer

When sampling multi-modal probability distributions, correctly estimating the relative probability of each mode, even when the modes have been discovered and locally sampled, remains challenging. We test a simple reweighting scheme designed…

Statistics Theory · Mathematics 2026-02-17 Pierre Monmarché

Optimum designs for parameter estimation in generalized regression models are standardly based on the Fisher information matrix (cf. Atkinson et al (2014) for a recent exposition). The corresponding optimality criteria are related to the…

Statistics Theory · Mathematics 2015-07-28 Katarína Burclová , Andrej Pázman

We investigate the separability properties of quantum states described by an extended Werner density matrix, where the classical component exhibits statistical dependence. By generalizing the classical part to allow correlations, we…

General Physics · Physics 2025-07-22 Toru Ohira

Information generating functions have been used for generating various entropy and divergence measures. In the present work, we introduce quantile based relative information generating function and study its properties. The proposed…

Statistics Theory · Mathematics 2024-12-04 Sankaran P. G. , Sunoj S. M. , Pavithra Hariharan

The quality of the inferences we make from pathogen sequence data is determined by the number and composition of pathogen sequences that make up the sample used to drive that inference. However, there remains limited guidance on how to best…

Populations and Evolution · Quantitative Biology 2023-06-13 Lucy D'Agostino McGowan , Shirlee Wohl , Justin Lessler

We characterize Martin-L\"of randomness and Schnorr randomness in terms of the merging of opinions, along the lines of the Blackwell-Dubins Theorem. After setting up a general framework for defining notions of merging randomness, we focus…

Logic · Mathematics 2026-03-10 Simon M. Huttegger , Sean Walsh , Francesca Zaffora Blando