English
Related papers

Related papers: Perturbed factor analysis: Accounting for group di…

200 papers

Mixtures of factor analysers (MFA) models represent a popular tool for finding structure in data, particularly high-dimensional data. While in most applications the number of clusters, and especially the number of latent factors within…

Methodology · Statistics 2023-07-17 Margarita Grushanina , Sylvia Frühwirth-Schnatter

In model-based clustering and classification, the cluster-weighted model constitutes a convenient approach when the random vector of interest constitutes a response variable Y and a set p of explanatory variables X. However, its…

Methodology · Statistics 2013-07-23 Sanjeena Subedi , Antonio Punzo , Salvatore Ingrassia , Paul D. McNicholas

Identification of features is a critical task in microbiome studies that is complicated by the fact that microbial data are high dimensional and heterogeneous. Masked by the complexity of the data, the problem of separating signals from…

Methodology · Statistics 2020-09-18 Liangliang Zhang , Yushu Shi , Kim-Anh Do , Christine B. Peterson , Robert R. Jenq

Joint Bayesian factor models are popular for characterizing relationships between multivariate correlated predictors and a response variable. Standard models assume that all variables, including both the predictors and the response, are…

Methodology · Statistics 2025-05-19 Glenn Palmer , David B. Dunson

The detrended fluctuation analysis (DFA) is one of the most widely used tools for the detection of long-range correlations in time series. Although DFA has found many interesting applications and has been shown as one of the best performing…

Statistical Mechanics · Physics 2020-03-18 G. Sikora , M. Hoell , A. Wylomanska , J. Gajda , A. V. Chechkin , H. Kantz

Principal stratification provides a causal inference framework for investigating treatment effects in the presence of a post-treatment variable. Principal strata play a key role in characterizing the treatment effect by identifying groups…

Standard linear modeling approaches make potentially simplistic assumptions regarding the structure of categorical effects that may obfuscate more complex relationships governing data. For example, recent work focused on the two-way…

Methodology · Statistics 2019-03-05 Thomas A. Metzger , Christopher T. Franck

Sub-visible particle analysis using flow imaging microscopy combined with deep learning has proven effective in identifying particle types, enabling the distinction of harmless components such as silicone oil from protein particles.…

Computer Vision and Pattern Recognition · Computer Science 2025-08-11 Utku Ozbulak , Michaela Cohrs , Hristo L. Svilenov , Joris Vankerschaver , Wesley De Neve

Matrix concentration inequalities and their recently discovered sharp counterparts provide powerful tools to bound the spectrum of random matrices whose entries are linear functions of independent random variables. However, in many…

Probability · Mathematics 2025-04-02 Afonso S. Bandeira , Kevin Lucca , Petar Nizić-Nikolac , Ramon van Handel

Federated Learning (FL) has emerged as an excellent solution for performing deep learning on different data owners without exchanging raw data. However, statistical heterogeneity in FL presents a key challenge, leading to a phenomenon of…

Computer Vision and Pattern Recognition · Computer Science 2025-04-14 Junfeng Liao , Sifan Wang , Ye Yuan , Riquan Zhang

We consider the design and analysis of multi-factor experiments using fractional factorial and incomplete designs within the potential outcome framework. These designs are particularly useful when limited resources make running a full…

Methodology · Statistics 2022-01-31 Nicole E. Pashley , Marie-Abele C. Bind

This paper studies a factor modeling-based approach for clustering high-dimensional data generated from a mixture of strongly correlated variables. Statistical modeling with correlated structures pervades modern applications in economics,…

Statistics Theory · Mathematics 2024-08-23 Shange Tang , Soham Jana , Jianqing Fan

The R-package phtt provides estimation procedures for panel data with large dimensions n, T, and general forms of unobservable heterogeneous effects. Particularly, the estimation procedures are those of Bai (2009) and Kneip, Sickles, and…

Computation · Statistics 2014-07-25 Oualid Bada , Dominik Liebl

Ever since the seminal work of R. A. Fisher and F. Yates, factorial designs have been an important experimental tool to simultaneously estimate the effects of multiple treatment factors. In factorial designs, the number of treatment…

Methodology · Statistics 2024-03-21 Lei Shi , Jingshen Wang , Peng Ding

Detrended fluctuation analysis (DFA) is a scaling analysis method used to estimate long-range power-law correlation exponents in noisy signals. Many noisy signals in real systems display trends, so that the scaling results obtained from the…

Data Analysis, Statistics and Probability · Physics 2009-11-07 Kun Hu , Plamen Ch. Ivanov , Zhi Chen , Pedro Carpena , H. Eugene Stanley

The ethical, social and legal issues surrounding facial analysis technologies have been widely debated in recent years. Key critics have argued that these technologies can perpetuate bias and discrimination, particularly against…

Computer Vision and Pattern Recognition · Computer Science 2025-02-11 Marco Rondina , Fabiana Vinci , Antonio Vetrò , Juan Carlos De Martin

The past 20 years have brought fundamental advances in modeling unobserved heterogeneity in panel data. Interactive Fixed Effects (IFE) proved to be a foundational framework, generalizing the standard one-way and two-way fixed effects…

Econometrics · Economics 2025-10-15 Jan Ditzen , Yiannis Karavias

In the context of machine learning, disparate impact refers to a form of systematic discrimination whereby the output distribution of a model depends on the value of a sensitive attribute (e.g., race or gender). In this paper, we propose an…

Information Theory · Computer Science 2018-05-14 Hao Wang , Berk Ustun , Flavio P. Calmon

The analysis of the interaction matrix between two distinct sets is essential across diverse fields, from pharmacovigilance to transcriptomics. Not all interactions are equally informative: a marker gene associated with a few specific…

Per- and polyfluoroalkyl substances (PFAS) are typically encountered as mixtures of distinct chemicals with distinct effects on multiple health outcomes. Estimating joint causal effects using spatially-dependent observed data is…

Methodology · Statistics 2026-03-18 Xiaodan Zhou , Brian J Reich , Shu Yang