中文
相关论文

相关论文: Nonparametric Bayesian Knockoff Generators for Fea…

200 篇论文

Feature selection is central to contemporary high-dimensional data analysis. Grouping structure among features arises naturally in various scientific problems. Many methods have been proposed to incorporate the grouping structure…

机器学习 · 计算机科学 2019-05-28 Guangyu Zhu , Tingting Zhao

Model-X knockoffs is a general procedure that can leverage any feature importance measure to produce a variable selection algorithm, which discovers true effects while rigorously controlling the number or fraction of false positives.…

统计方法学 · 统计学 2020-12-07 Zhimei Ren , Yuting Wei , Emmanuel Candès

Controlling the False Discovery Rate (FDR) in a variable selection procedure is critical for reproducible discoveries, and it has been extensively studied in sparse linear models. However, it remains largely open in scenarios where the…

统计方法学 · 统计学 2023-11-16 Yang Cao , Xinwei Sun , Yuan Yao

Knockoff variable selection is a powerful framework that creates synthetic knockoff variables to mirror the correlation structure of the observed features, enabling principled control of the false discovery rate in variable selection.…

统计方法学 · 统计学 2025-08-21 Evan Mason , Zhe Fei

Many contemporary large-scale applications involve building interpretable models linking a large set of potential covariates to a response in a nonlinear fashion, such as when the response is binary. Although this modeling problem has been…

统计方法学 · 统计学 2017-12-13 Emmanuel Candes , Yingying Fan , Lucas Janson , Jinchi Lv

Model-free knockoffs is a recently proposed technique for identifying covariates that is likely to have an effect on a response variable. The method is an efficient method to control the false discovery rate in hypothesis tests for separate…

统计方法学 · 统计学 2019-03-29 Lars Holden , Kristoffer Hellton

Controlled feature selection aims to discover the features a response depends on while limiting the false discovery rate (FDR) to a predefined level. Recently, multiple deep-learning-based methods have been proposed to perform controlled…

机器学习 · 统计学 2022-10-24 Derek Hansen , Brian Manzo , Jeffrey Regier

We consider the variable selection problem, which seeks to identify important variables influencing a response $Y$ out of many candidate features $X_1, \ldots, X_p$. We wish to do so while offering finite-sample guarantees about the…

统计方法学 · 统计学 2019-02-12 Rina Foygel Barber , Emmanuel J. Candès , Richard J. Samworth

The knockoff filter introduced by Barber and Cand\`es 2016 is an elegant framework for controlling the false discovery rate in variable selection. While empirical results indicate that this methodology is not too conservative, there is no…

统计理论 · 数学 2020-01-13 Jingbo Liu , Philippe Rigollet

Model-X knockoffs is a flexible wrapper method for high-dimensional regression algorithms, which provides guaranteed control of the false discovery rate (FDR). Due to the randomness inherent to the method, different runs of model-X…

统计方法学 · 统计学 2023-09-01 Zhimei Ren , Rina Foygel Barber

The fixed-X knockoff filter is a flexible framework for variable selection with false discovery rate (FDR) control in linear models with arbitrary design matrices (of full column rank) and it allows for finite-sample selective inference via…

统计理论 · 数学 2023-11-28 Mehrdad Pournaderi , Yu Xiang

The rapid generation of complex, highly skewed, and zero-inflated multi-source count data poses significant challenges for variable selection, particularly in biomedical domains like tumor development and metabolic dysregulation. To address…

应用统计 · 统计学 2025-11-11 Shan Tang , Shanjun Mao , Shourong Ma , Falong Tan

We describe a series of algorithms that efficiently implement Gaussian model-X knockoffs to control the false discovery rate on large scale feature selection problems. Identifying the knockoff distribution requires solving a large scale…

机器学习 · 计算机科学 2020-06-17 Armin Askari , Quentin Rebjock , Alexandre d'Aspremont , Laurent El Ghaoui

Model-X knockoffs is a wrapper that transforms essentially any feature importance measure into a variable selection algorithm, which discovers true effects while rigorously controlling the expected fraction of false positives. A frequently…

统计方法学 · 统计学 2024-03-12 Stephen Bates , Emmanuel Candès , Lucas Janson , Wenshuo Wang

The goal of feature selection is to identify important features that are relevant to explain an outcome variable. Most of the work in this domain has focused on identifying globally relevant features, which are features that are related to…

机器学习 · 统计学 2019-05-30 Jaime Roquero Gimenez , James Zou

Knockoffs provide a general framework for controlling the false discovery rate when performing variable selection. Much of the Knockoffs literature focuses on theoretical challenges and we recognize a need for bringing some of the current…

统计方法学 · 统计学 2020-10-28 Matthias Kormaksson , Luke J. Kelly , Xuan Zhu , Sibylle Haemmerle , Luminita Pricop , David Ohlssen

Generative Bayesian Filtering (GBF) provides a powerful and flexible framework for performing posterior inference in complex nonlinear and non-Gaussian state-space models. Our approach extends Generative Bayesian Computation (GBC) to…

统计方法学 · 统计学 2025-11-07 Edoardo Marcelli , Sean O'Hagan , Veronika Rockova

We investigate the robustness of the model-X knockoffs framework with respect to the misspecified or estimated feature distribution. We achieve such a goal by theoretically studying the feature selection performance of a practically…

统计方法学 · 统计学 2024-06-06 Yingying Fan , Lan Gao , Jinchi Lv

In high dimensional variable selection problems, statisticians often seek to design multiple testing procedures that control the False Discovery Rate (FDR), while concurrently identifying a greater number of relevant variables. Model-X…

统计理论 · 数学 2023-07-25 Taejoo Ahn , Licong Lin , Song Mei

It is very challenging to select informative features from tens of thousands of measured features in high-throughput data analysis. Recently, several parametric/regression models have been developed utilizing the gene network information to…

应用统计 · 统计学 2014-08-01 Yize Zhao , Jian Kang , Tianwei Yu