English
Related papers

Related papers: Comment on "mbtransfer: Microbiome intervention an…

200 papers

Missing data represents a fundamental challenge in machine learning applications, often reducing model performance and reliability. This problem is particularly acute in fields like bioinformatics and clinical machine learning, where…

Machine Learning · Computer Science 2025-09-04 Fatemeh Azad , Zoran Bosnić , Matjaž Kukar

In many experimental contexts, whether and how network interactions impact the outcome of interest for both treated and untreated individuals are key concerns. Networks data is often assumed to perfectly represent these possible…

Methodology · Statistics 2024-03-12 Morgan Hardy , Rachel M. Heath , Wesley Lee , Tyler H. McCormick

This paper focuses on the influence of a misspecified covariance structure on false discovery rate for the large scale multiple testing problem. Specifically, we evaluate the influence on the marginal distribution of local fdr statistics,…

Statistics Theory · Mathematics 2019-02-19 Ye Liang , Joshua D. Habiger , Xiaoyi Min

It is common practice in using regression type models for inferring causal effects, that inferring the correct causal relationship requires extra covariates are included or ``adjusted for''. Without performing this adjustment erroneous…

Machine Learning · Statistics 2019-06-18 David Rohde

Many statistical models have high accuracy on test benchmarks, but are not explainable, struggle in low-resource scenarios, cannot be reused for multiple tasks, and cannot easily integrate domain expertise. These factors limit their use,…

Computation and Language · Computer Science 2021-09-29 Andrew Lee , Jonathan K. Kummerfeld , Lawrence C. An , Rada Mihalcea

Algorithmic decision-making systems sometimes produce errors or skewed predictions toward a particular group, leading to unfair results. Debiasing practices, applied at different stages of the development of such systems, occasionally…

Artificial Intelligence · Computer Science 2025-05-26 Juliett Suárez Ferreira , Marija Slavkovik , Jorge Casillas

Microbiota contribute to many dimensions of host phenotype, including disease. To link specific microbes to specific phenotypes, microbiome-wide association studies compare microbial abundances between two groups of samples. Abundance…

Applications · Statistics 2018-01-31 Rajita Menon , Vivek Ramanan , Kirill S. Korolev

Online experiments (A/B tests) are widely regarded as the gold standard for evaluating recommender system variants and guiding launch decisions. However, a variety of biases can distort the results of the experiment and mislead…

Information Retrieval · Computer Science 2025-09-03 Chen Zheng , Zhenyu Zhao

Percentiles have been established in bibliometrics as an important alternative to mean-based indicators for obtaining a normalized citation impact of publications. Percentiles have a number of advantages over standard bibliometric…

Digital Libraries · Computer Science 2012-11-05 Lutz Bornmann , Loet Leydesdorff , Ruediger Mutz

We address the common yet often-overlooked selection bias in interventional studies, where subjects are selectively enrolled into experiments. For instance, participants in a drug trial are usually patients of the relevant disease; A/B…

Machine Learning · Computer Science 2025-03-11 Haoyue Dai , Ignavier Ng , Jianle Sun , Zeyu Tang , Gongxu Luo , Xinshuai Dong , Peter Spirtes , Kun Zhang

Diffuse Reflectance Spectroscopy has demonstrated a strong aptitude for identifying and differentiating biological tissues. However, the broadband and smooth nature of these signals require algorithmic processing, as they are often…

Image and Video Processing · Electrical Eng. & Systems 2025-03-06 Nicola Rossberg , Celina L. Li , Simone Innocente , Stefan Andersson-Engels , Katarzyna Komolibus , Barry O'Sullivan , Andrea Visentin

I study identification, estimation and inference for spillover effects in experiments where units' outcomes may depend on the treatment assignments of other units within a group. I show that the commonly-used reduced-form linear-in-means…

Econometrics · Economics 2022-01-21 Gonzalo Vazquez-Bare

Imputation methods play a critical role in enhancing the quality of practical time-series data, which often suffer from pervasive missing values. Recently, diffusion-based generative imputation methods have demonstrated remarkable success…

Machine Learning · Computer Science 2025-10-03 Zeqi Ye , Minshuo Chen

This work considers Bayesian inference under misspecification for complex statistical models comprised of simpler submodels, referred to as modules, that are coupled together. Such ``multi-modular" models often arise when combining…

Statistics Theory · Mathematics 2023-08-02 David T. Frazier , David J. Nott

Untethered mobile milli/microrobots hold transformative potential for interventional medicine by enabling more precise and entirely non-invasive diagnosis and therapy. Realizing this promise requires bridging the gap between groundbreaking…

Robotics · Computer Science 2025-10-15 Hakan Ceylan , Edoardo Sinibaldi , Sanjay Misra , Pankaj J. Pasricha , Dietmar W. Hutmacher

Due to complex experimental settings, missing values are common in biomedical data. To handle this issue, many methods have been proposed, from ignoring incomplete instances to various data imputation approaches. With the recent rise of…

Machine Learning · Computer Science 2020-05-14 Kristian Miok , Dong Nguyen-Doan , Marko Robnik-Šikonja , Daniela Zaharie

We present a new approach to semiparametric inference using corrected posterior distributions. The method allows us to leverage the adaptivity, regularization and predictive power of nonparametric Bayesian procedures to estimate…

Methodology · Statistics 2023-06-21 Andrew Yiu , Edwin Fong , Chris Holmes , Judith Rousseau

Linear transformation model provides a general framework for analyzing censored survival data with covariates. The proportional hazards and proportional odds models are special cases of the linear transformation model. In biomedical…

Methodology · Statistics 2022-05-11 Sudheesh K. K. , Deemat C. Mathew , Litty Mathew , Min Xie

Many observational studies feature irregular longitudinal data, where the observation times are not common across individuals in the study. Further, the observation times may be related to the longitudinal outcome. In this setting, failing…

Methodology · Statistics 2024-05-27 Grace Tompkins , Joel A Dubin , Michael Wallace

Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on predictions made by these models is crucial for improving their trustworthiness. Current…

Machine Learning · Computer Science 2024-06-21 Myeongseob Ko , Feiyang Kang , Weiyan Shi , Ming Jin , Zhou Yu , Ruoxi Jia