中文
相关论文

相关论文: Multilayer Knockoff Filter: Controlled variable se…

200 篇论文

Competition-based FDR control has been commonly used for over a decade in the computational mass spectrometry community (Elias and Gygi, 2007). Recently, the approach has gained significant popularity in other fields after Barber and Candes…

统计方法学 · 统计学 2019-11-14 Kristen Emery , Syamand Hasam , William Stafford Noble , Uri Keich

Model-X knockoffs allows analysts to perform feature selection using almost any machine learning algorithm while still provably controlling the expected proportion of false discoveries. To apply model-X knockoffs, one must construct…

统计方法学 · 统计学 2021-06-30 Asher Spector , Lucas Janson

In many scientific problems, researchers try to relate a response variable $Y$ to a set of potential explanatory variables $X = (X_1,\dots,X_p)$, and start by trying to identify variables that contribute to this relationship. In statistical…

统计理论 · 数学 2020-10-07 Wenshuo Wang , Lucas Janson

Knockoffs provide a general framework for controlling the false discovery rate when performing variable selection. Much of the Knockoffs literature focuses on theoretical challenges and we recognize a need for bringing some of the current…

统计方法学 · 统计学 2020-10-28 Matthias Kormaksson , Luke J. Kelly , Xuan Zhu , Sibylle Haemmerle , Luminita Pricop , David Ohlssen

Feature Selection (FS) plays an important role in learning and classification tasks. The object of FS is to select the relevant and non-redundant features. Considering the huge amount number of features in real-world applications, FS…

机器学习 · 计算机科学 2019-06-19 Fatma BenSaid , Adel M. Alimi

This paper proposes a reliable neural network pruning algorithm by setting up a scientific control. Existing pruning methods have developed various hypotheses to approximate the importance of filters to the network and then execute filter…

计算机视觉与模式识别 · 计算机科学 2021-01-12 Yehui Tang , Yunhe Wang , Yixing Xu , Dacheng Tao , Chunjing Xu , Chao Xu , Chang Xu

Controlling false discovery rate (FDR) while leveraging the side information of multiple hypothesis testing is an emerging research topic in modern data science. Existing methods rely on the test-level covariates while ignoring possible…

机器学习 · 统计学 2021-01-26 Lin Qiu , Nils Murrugarra-Llerena , Vítor Silva , Lin Lin , Vernon M. Chinchilli

Bilateral filtering (BF) is one of the most classical denoising filters, however, the manually initialized filtering kernel hampers its adaptivity across images with various characteristics. To deal with image variation (i.e.,…

计算机视觉与模式识别 · 计算机科学 2019-12-24 Feihong Liu , Jun Feng , Pew-Thian Yap , Dinggang Shen

Knockoffs is a new framework for controlling the false discovery rate (FDR) in multiple hypothesis testing problems involving complex statistical models. While there has been great emphasis on Type-I error control, Type-II errors have been…

统计方法学 · 统计学 2017-12-19 Asaf Weinstein , Rina Barber , Emmanuel Candes

Genome-wide association studies (GWAS) often find association signals between many genetic variants and traits of interest in a genomic region. Functional annotations of these variants provide valuable prior information that helps…

统计方法学 · 统计学 2026-01-07 Xiangyu Zhang , Lijun Wang , Changjun Li , Chen Lin , Hongyu Zhao

This paper introduces a machine for sampling approximate model-X knockoffs for arbitrary and unspecified data distributions using deep generative models. The main idea is to iteratively refine a knockoff sampling mechanism until a criterion…

统计方法学 · 统计学 2020-03-03 Yaniv Romano , Matteo Sesia , Emmanuel J. Candès

Conditional testing via the knockoff framework allows one to identify -- among large number of possible explanatory variables -- those that carry unique information about an outcome of interest, and also provides a false discovery rate…

统计方法学 · 统计学 2024-03-05 Benjamin B Chu , Jiaqi Gu , Zhaomeng Chen , Tim Morrison , Emmanuel Candes , Zihuai He , Chiara Sabatti

Identifying which variables do influence a response while controlling false positives pervades statistics and data science. In this paper, we consider a scenario in which we only have access to summary statistics, such as the values of…

统计方法学 · 统计学 2024-02-21 Zhaomeng Chen , Zihuai He , Benjamin B. Chu , Jiaqi Gu , Tim Morrison , Chiara Sabatti , Emmanuel Candès

We introduce DiffKnock, a diffusion-based knockoff framework for high-dimensional feature selection with finite-sample false discovery rate (FDR) control. DiffKnock addresses two key limitations of existing knockoff methods: preserving…

统计方法学 · 统计学 2025-10-03 Heng Ge , Qing Lu

We introduce local conditional hypotheses that express how the relation between explanatory variables and outcomes changes across different contexts, described by covariates. By expanding upon the model-X knockoff filter, we show how to…

统计方法学 · 统计学 2026-01-12 Paula Gablenz , Matteo Sesia , Tianshu Sun , Chiara Sabatti

We introduce a novel privatization framework for high-dimensional controlled variable selection. Our framework enables rigorous False Discovery Rate (FDR) control under differential privacy constraints. While the Model-X knockoff procedure…

机器学习 · 统计学 2025-08-08 Yuxuan Tao , Adel Javanmard

Despite the popularity of feature importance (FI) measures in interpretable machine learning, the statistical adequacy of these methods is rarely discussed. From a statistical perspective, a major distinction is between analyzing a…

机器学习 · 统计学 2023-05-03 Kristin Blesch , David S. Watson , Marvin N. Wright

Predictive modeling often uses black box machine learning methods, such as deep neural networks, to achieve state-of-the-art performance. In scientific domains, the scientist often wishes to discover which features are actually important…

机器学习 · 统计学 2020-08-03 Mukund Sudarshan , Wesley Tansey , Rajesh Ranganath

Feature selection, which is a technique to select key features in recommender systems, has received increasing research attention. Recently, Adaptive Feature Selection (AdaFS) has shown remarkable performance by adaptively selecting…

信息检索 · 计算机科学 2023-09-07 Youngjune Lee , Yeongjong Jeong , Keunchan Park , SeongKu Kang

We develop a flexible feature selection framework based on deep neural networks that approximately controls the false discovery rate (FDR), a measure of Type-I error. The method applies to architectures whose first layer is fully connected.…

机器学习 · 统计学 2026-02-10 Kazuma Sawaya