中文
相关论文

相关论文: Significance Analysis for Pairwise Variable Select…

200 篇论文

There is an increasing interest in estimating heterogeneity in causal effects in randomized and observational studies. However, little research has been conducted to understand heterogeneity in an instrumental variables study. In this work,…

统计方法学 · 统计学 2021-01-20 Michael Johnson , Jiongyi Cao , Hyunseung Kang

Alcohol misuse is a key target of public health strategies aimed at reducing cardiovascular risk. The effect of excessive alcohol consumption on blood pressure may vary systematically with individuals' unobserved propensity to engage in…

统计方法学 · 统计学 2026-03-11 Ashish Patel , Francis J DiTraglia , Stephen Burgess

The large-scale multiple testing inherent to high throughput biological data necessitates very high statistical stringency and thus true effects in data are difficult to detect unless they have high effect sizes. One solution to this…

统计方法学 · 统计学 2017-12-21 Mohamad S. Hasan

Substantial progress has been made in identifying single genetic variants predisposing to common complex diseases. Nonetheless, the genetic etiology of human diseases remains largely unknown. Human complex diseases are likely influenced by…

统计方法学 · 统计学 2014-05-27 Zihuai He , Min Zhang , Xiaowei Zhan , Qing Lu

Practitioners use feature importance to rank and eliminate weak predictors during model development in an effort to simplify models and improve generality. Unfortunately, they also routinely conflate such feature importance measures with…

机器学习 · 计算机科学 2020-06-09 Terence Parr , James D. Wilson , Jeff Hamrick

A common approach for feature selection is to examine the variable importance scores for a machine learning model, as a way to understand which features are the most relevant for making predictions. Given the significance of feature…

机器学习 · 计算机科学 2021-05-13 Jack Dunn , Luca Mingardi , Ying Daisy Zhuo

A sequential importance sampling algorithm is developed for the distribution that results when a matrix of independent, but not identically distributed, Bernoulli random variables is conditioned on a given sequence of row and column sums.…

统计计算 · 统计学 2013-01-18 Matthew T. Harrison , Jeffrey W. Miller

The issue of variance components testing arises naturally when building mixed-effects models, to decide which effects should be modeled as fixed or random. While tests for fixed effects are available in R for models fitted with lme4, tools…

统计方法学 · 统计学 2021-05-18 Charlotte Baey , Estelle Kuhn

Popular statistical software provides Bayesian information criterion (BIC) for multilevel models or linear mixed models. However, it has been observed that the combination of statistical literature and software documentation has led to…

统计方法学 · 统计学 2022-06-24 Sun-Joo Cho , Hao Wu , Matthew Naveiras

Causal inference on populations embedded in social networks poses technical challenges, since the typical no interference assumption frequently does not hold. Existing methods developed in the context of network interference rely upon the…

统计方法学 · 统计学 2024-04-12 Vanessa McNealis , Erica E. M. Moodie , Nema Dean

In modern biomedical research, it is ubiquitous to have multiple data sets measured on the same set of samples from different views (i.e., multi-view data). For example, in genetic studies, multiple genomic data sets at different molecular…

统计方法学 · 统计学 2017-03-20 Gen Li , Sungkyu Jung

In many scientific problems, researchers try to relate a response variable $Y$ to a set of potential explanatory variables $X = (X_1,\dots,X_p)$, and start by trying to identify variables that contribute to this relationship. In statistical…

统计理论 · 数学 2020-10-07 Wenshuo Wang , Lucas Janson

In this paper, we introduce Partial Information Decomposition of Features (PIDF), a new paradigm for simultaneous data interpretability and feature selection. Contrary to traditional methods that assign a single importance value, our…

机器学习 · 计算机科学 2025-11-17 Charles Westphal , Stephen Hailes , Mirco Musolesi

Variable selection is fundamental to high-dimensional statistical modeling. Many variable selection techniques may be implemented by maximum penalized likelihood using various penalty functions. Optimizing the penalized likelihood function…

统计理论 · 数学 2007-06-13 David R. Hunter , Runze Li

Mixed linear models are commonly used in repeated measures studies. They account for the dependence amongst observations obtained from the same experimental unit. Oftentimes, the number of observations is small, and it is thus important to…

统计方法学 · 统计学 2011-08-05 Tatiane F. N. Melo , Silvia L. P. Ferrari , Francisco Cribari-Neto

Propensity score matching is commonly used to draw causal inference from observational survival data. However, its asymptotic properties have yet to be established, and variance estimation is still open to debate. We derive the statistical…

统计方法学 · 统计学 2024-12-24 Tongrong Wang , Honghe Zhao , Shu Yang , Shuhan Tang , Zhanglin Cui , Li Li , Douglas E. Faries

Propensity Score Matching (PSM) is a causal inference technique that is used as a substitution for experimental methods when it is not possible to implement them due to logistical and ethical concerns. By using a logistic classifier to…

应用统计 · 统计学 2025-01-15 Elizabeth Mohney , Alexey Shvets

Multivariate meta-analysis is gaining prominence in evidence synthesis research because it enables simultaneous synthesis of multiple correlated outcome data, and random-effects models have generally been used for addressing between-studies…

统计方法学 · 统计学 2021-07-14 Hisashi Noma , Kengo Nagashima , Toshi A. Furukawa

In the evaluation of treatment effects, it is of major policy interest to know if the treatment is beneficial for some and harmful for others, a phenomenon known as qualitative interaction. We formulate this question as a multiple testing…

统计方法学 · 统计学 2017-08-29 Qingyuan Zhao , Dylan S. Small , Weijie Su

Estimating the strength of dependency between two variables is fundamental for exploratory analysis and many other applications in data mining. For example: non-linear dependencies between two continuous variables can be explored with the…

机器学习 · 统计学 2016-01-21 Simone Romano , Nguyen Xuan Vinh , James Bailey , Karin Verspoor