English
Related papers

Related papers: Fisher exact scanning for dependency

200 papers

The Gaussian graphical model (GGM) incorporates an undirected graph to represent the conditional dependence between variables, with the precision matrix encoding partial correlation between pair of variables given the others. To achieve…

Methodology · Statistics 2023-07-03 Yueqi Qian , Xianghong Hu , Can Yang

The flexibility and wide applicability of the Fisher randomization test (FRT) makes it an attractive tool for assessment of causal effects of interventions from modern-day randomized experiments that are increasing in size and complexity.…

Methodology · Statistics 2020-04-21 Xiaokang Luo , Tirthankar Dasgupta , Minge Xie , Regina Liu

The exact estimation of latent variable models with big data is known to be challenging. The latents have to be integrated out numerically, and the dimension of the latent variables increases with the sample size. This paper develops a…

Econometrics · Economics 2023-06-27 Ruben Loaiza-Maya , Didier Nibbering , Dan Zhu

We discuss certain basic features of the equation-free (EF) approach to modeling and computation for complex/multiscale systems. We focus on links between the equation-free approach and tools from systems and control theory (design of…

Cellular Automata and Lattice Gases · Physics 2007-05-23 C. I. Siettos , R. Rico-Martinez , I. G. kevrekidis

Frequent Subgraph Mining (FSM) is the process of identifying common subgraph patterns that surpass a predefined frequency threshold. While FSM is widely applicable in fields like bioinformatics, chemical analysis, and social network anomaly…

Databases · Computer Science 2024-04-03 Akshit Sharma , Sam Reinher , Dinesh Mehta , Bo Wu

The expectation-maximization (EM) algorithm is an iterative computational method to calculate the maximum likelihood estimators (MLEs) from the sample data. It converts a complicated one-time calculation for the MLE of the incomplete data…

Computation · Statistics 2016-08-08 Lingyao Meng

Linear mixed-effect models with two variance components are often used when variability comes from two sources. In genetics applications, variation in observed traits can be attributed to biological and environmental effects, and the…

Methodology · Statistics 2015-01-19 Qianshun Cheng , Xu Gao , Ryan Martin

Software vulnerability detection can be formulated as a binary classification problem that determines whether a given code snippet contains security defects. Existing multimodal methods typically fuse Natural Code Sequence (NCS)…

Software Engineering · Computer Science 2026-04-24 Yun Bian , Yi Chen , HaiQuan Wang , ShiHao Li , Zhe Cui

Graphical models describe associations between variables through the notion of conditional independence. Gaussian graphical models are a widely used class of such models where the relationships are formalized by non-null entries of the…

Methodology · Statistics 2023-08-08 Sagnik Bhadury , Riten Mitra , Jeremy T. Gaskins

Accurate pest population monitoring and tracking their dynamic changes are crucial for precision agriculture decision-making. A common limitation in existing vision-based automatic pest counting research is that models are typically…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Xumin Gao , Mark Stevens , Grzegorz Cielniak

In this note we present a fully information theoretic approach to renormalization inspired by Bayesian statistical inference, which we refer to as Bayesian Renormalization. The main insight of Bayesian Renormalization is that the Fisher…

High Energy Physics - Theory · Physics 2023-10-11 David S. Berman , Marc S. Klinger , Alexander G. Stapleton

Starting from the Fisher matrix for counts in cells, I derive the full Fisher matrix for surveys of multiple tracers of large-scale structure. The key assumption is that the inverse of the covariance of the galaxy counts is given by the…

Cosmology and Nongalactic Astrophysics · Physics 2012-05-28 L. Raul Abramo

Bayesian sparse factor models have proven useful for characterizing dependence in multivariate data, but scaling computation to large numbers of samples and dimensions is problematic. We propose expandable factor analysis for scalable…

Methodology · Statistics 2018-06-21 Sanvesh Srivastava , Barbara E. Engelhardt , David B. Dunson

We propose a simple single-step multiple testing procedure that asymptotically controls the family-wise error rate (FWER) at the desired level exactly under the equicorrelated multivariate Gaussian setup. The method is shown to be…

Statistics Theory · Mathematics 2025-08-14 Swarnadeep Datta , Monitirtha Dey

Studying phenotype-gene association can uncover mechanism of diseases and develop efficient treatments. In complex disease where multiple phenotypes are available and correlated, analyzing and interpreting associated genes for each…

Methodology · Statistics 2021-12-14 Yujia Li , Yusi Fang , Peng Liu , George C. Tseng

We show how to calculate individual terms of the Edgeworth series to approximate the distribution of the Pearson correlation coefficient with the help of a simple Mathematica program. We also demonstrate how to eliminate the corresponding…

Statistics Theory · Mathematics 2022-08-11 Jan Vrbik

The applications of traditional statistical feature selection methods to high-dimension, low sample-size data often struggle and encounter challenging problems, such as overfitting, curse of dimensionality, computational infeasibility, and…

Machine Learning · Statistics 2023-12-19 Kexuan Li , Fangfang Wang , Lingli Yang , Ruiqi Liu

In this paper we propose a multiscale scanning method to determine active components of a quantity $f$ w.r.t. a dictionary $\mathcal{U}$ from observations $Y$ in an inverse regression model $Y=Tf+\xi$ with linear operator $T$ and general…

Methodology · Statistics 2017-06-28 Katharina Proksch , Frank Werner , Axel Munk

The Fisher-Snedecor $\mathcal{F}$ distribution has been recently proposed as a more accurate and mathematically tractable composite fading model than traditional established models in some practical cases. In this paper, we firstly derive…

Information Theory · Computer Science 2019-11-27 Hongyang Du , Jiayi Zhang , Kostas P. Peppas , Hui Zhao , Bo Ai , Xiaodan Zhang

Testing for independence between two random vectors is a fundamental problem in statistics. It is observed from empirical studies that many existing omnibus consistent tests may not work well for some strongly nonmonotonic and nonlinear…

Methodology · Statistics 2024-02-27 Kai Xu , Yeqing Zhou , Liping Zhu , Runze Li