中文
相关论文

相关论文: Powerful Knockoffs via Minimizing Reconstructabili…

200 篇论文

Model-X knockoffs is a wrapper that transforms essentially any feature importance measure into a variable selection algorithm, which discovers true effects while rigorously controlling the expected fraction of false positives. A frequently…

统计方法学 · 统计学 2024-03-12 Stephen Bates , Emmanuel Candès , Lucas Janson , Wenshuo Wang

We consider the variable selection problem, which seeks to identify important variables influencing a response $Y$ out of many candidate features $X_1, \ldots, X_p$. We wish to do so while offering finite-sample guarantees about the…

统计方法学 · 统计学 2019-02-12 Rina Foygel Barber , Emmanuel J. Candès , Richard J. Samworth

This paper introduces a machine for sampling approximate model-X knockoffs for arbitrary and unspecified data distributions using deep generative models. The main idea is to iteratively refine a knockoff sampling mechanism until a criterion…

统计方法学 · 统计学 2020-03-03 Yaniv Romano , Matteo Sesia , Emmanuel J. Candès

Many contemporary large-scale applications involve building interpretable models linking a large set of potential covariates to a response in a nonlinear fashion, such as when the response is binary. Although this modeling problem has been…

统计方法学 · 统计学 2017-12-13 Emmanuel Candes , Yingying Fan , Lucas Janson , Jinchi Lv

An important problem in machine learning and statistics is to identify features that causally affect the outcome. This is often impossible to do from purely observational data, and a natural relaxation is to identify features that are…

机器学习 · 统计学 2019-05-30 Jaime Roquero Gimenez , Amirata Ghorbani , James Zou

Model-X knockoffs is a general procedure that can leverage any feature importance measure to produce a variable selection algorithm, which discovers true effects while rigorously controlling the number or fraction of false positives.…

统计方法学 · 统计学 2020-12-07 Zhimei Ren , Yuting Wei , Emmanuel Candès

The Model-X knockoff procedure has recently emerged as a powerful approach for feature selection with statistical guarantees. The advantage of knockoff is that if we have a good model of the features X, then we can identify salient features…

机器学习 · 统计学 2019-05-30 Jaime Roquero Gimenez , James Zou

As new Model-X knockoff construction techniques are developed, primarily concerned with determining the correct conditional distribution from which to sample, we focus less on deriving the correct multivariate distribution and instead ask…

统计方法学 · 统计学 2025-02-05 Christopher Hemmens , Stephan Robert-Nicoud

We investigate the robustness of the model-X knockoffs framework with respect to the misspecified or estimated feature distribution. We achieve such a goal by theoretically studying the feature selection performance of a practically…

统计方法学 · 统计学 2024-06-06 Yingying Fan , Lan Gao , Jinchi Lv

Recently, the scheme of model-X knockoffs was proposed as a promising solution to address controlled feature selection under high-dimensional finite-sample settings. However, the procedure of model-X knockoffs depends heavily on the…

统计方法学 · 统计学 2022-03-10 Xuebin Zhao , Hong Chen , Yingjie Wang , Weifu Li , Tieliang Gong , Yulong Wang , Feng Zheng

Variable selection properties of procedures utilizing penalized-likelihood estimates is a central topic in the study of high dimensional linear regression problems. Existing literature emphasizes the quality of ranking of the variables by…

Model-X knockoff has garnered significant attention among various feature selection methods due to its guarantees for controlling the false discovery rate (FDR). Since its introduction in parametric design, knockoff techniques have evolved…

机器学习 · 计算机科学 2024-11-11 Hongyu Shen , Yici Yan , Zhizhen Zhao

In feature selection problems, knockoffs are synthetic controls for the original features. Employing knockoffs allows analysts to use nearly any variable importance measure or "feature statistic" to select features while rigorously…

统计方法学 · 统计学 2024-10-02 Asher Spector , William Fithian

One limitation of the most statistical/machine learning-based variable selection approaches is their inability to control the false selections. A recently introduced framework, model-x knockoffs, provides that to a wide range of models but…

机器学习 · 统计学 2025-09-03 Deniz Koyuncu , Alex Gittens , Bülent Yener

Model-free knockoffs is a recently proposed technique for identifying covariates that is likely to have an effect on a response variable. The method is an efficient method to control the false discovery rate in hypothesis tests for separate…

统计方法学 · 统计学 2019-03-29 Lars Holden , Kristoffer Hellton

Controlling the False Discovery Rate (FDR) is critical for reproducible variable selection, especially given the prevalence of complex predictive modeling. The recent Split Knockoff method, an extension of the canonical Knockoffs framework,…

统计方法学 · 统计学 2025-09-05 Yang Cao , Hangyu Lin , Xinwei Sun , Yuan Yao

Consider a case-control study in which we have a random sample, constructed in such a way that the proportion of cases in our sample is different from that in the general population---for instance, the sample is constructed to achieve a…

统计方法学 · 统计学 2019-01-01 Rina Foygel Barber , Emmanuel Candes

The knockoffs is a recently proposed powerful framework that effectively controls the false discovery rate (FDR) for variable selection. However, none of the existing knockoff solutions are directly suited to handle multivariate or…

统计方法学 · 统计学 2024-06-28 Xinghao Qiao , Mingya Long , Qizhai Li

We describe a series of algorithms that efficiently implement Gaussian model-X knockoffs to control the false discovery rate on large scale feature selection problems. Identifying the knockoff distribution requires solving a large scale…

机器学习 · 计算机科学 2020-06-17 Armin Askari , Quentin Rebjock , Alexandre d'Aspremont , Laurent El Ghaoui

A core strength of knockoff methods is their virtually limitless customizability, allowing an analyst to exploit machine learning algorithms and domain knowledge without threatening the method's robust finite-sample false discovery rate…

统计理论 · 数学 2021-07-15 Xiao Li , William Fithian
‹ 上一页 1 2 3 10 下一页 ›