中文
相关论文

相关论文: Powerful Knockoffs via Minimizing Reconstructabili…

200 篇论文

Modern scientific studies often require the identification of a subset of relevant explanatory variables, in the attempt to understand an interesting phenomenon. Several statistical methods have been developed to automate this task, but…

统计方法学 · 统计学 2019-05-14 Matteo Sesia , Chiara Sabatti , Emmanuel J. Candès

Feature selection prepares the AI-readiness of data by eliminating redundant features. Prior research falls into two primary categories: i) Supervised Feature Selection, which identifies the optimal feature subset based on their relevance…

机器学习 · 计算机科学 2024-03-08 Xinyuan Wang , Dongjie Wang , Wangyang Ying , Rui Xie , Haifeng Chen , Yanjie Fu

In many fields of science, we observe a response variable together with a large number of potential explanatory variables, and would like to be able to discover which variables are truly associated with the response. At the same time, we…

统计方法学 · 统计学 2015-10-15 Rina Foygel Barber , Emmanuel J. Candès

We apply the knockoff procedure to factor selection in finance. By building fake but realistic factors, this procedure makes it possible to control the fraction of false discovery in a given set of factors. To show its versatility, we apply…

统计金融 · 定量金融 2021-07-07 Damien Challet , Christian Bongiorno , Guillaume Pelletier

Model-X knockoffs is a flexible wrapper method for high-dimensional regression algorithms, which provides guaranteed control of the false discovery rate (FDR). Due to the randomness inherent to the method, different runs of model-X…

统计方法学 · 统计学 2023-09-01 Zhimei Ren , Rina Foygel Barber

The model-X conditional randomization test is a generic framework for conditional independence testing, unlocking new possibilities to discover features that are conditionally associated with a response of interest while controlling type-I…

机器学习 · 计算机科学 2023-02-21 Shalev Shaer , Yaniv Romano

Vine copulas are a flexible tool for high-dimensional dependence modeling. In this article, we discuss the generation of approximate model-X knockoffs with vine copulas. It is shown how Gaussian knockoffs can be generalized to Gaussian…

统计方法学 · 统计学 2022-10-21 Malte S. Kurz

The knockoff filter of Barber and Candes (arXiv:1404.5609) is a flexible framework for multiple testing in supervised learning models, based on introducing synthetic predictor variables to control the false discovery rate (FDR). Using the…

统计方法学 · 统计学 2024-11-26 Yixiang Luo , William Fithian , Lihua Lei

We tackle the problem of selecting from among a large number of variables those that are 'important' for an outcome. We consider situations where groups of variables are also of interest in their own right. For example, each variable might…

统计方法学 · 统计学 2018-08-13 Eugene Katsevich , Chiara Sabatti

This paper develops a framework for testing for associations in a possibly high-dimensional linear model where the number of features/variables may far exceed the number of observational units. In this framework, the observations are split…

统计方法学 · 统计学 2018-05-04 Rina Foygel Barber , Emmanuel J. Candes

The traditional framework for feature selection treats all features as costing the same amount. However, in reality, a scientist often has considerable discretion regarding which variables to measure, and the decision involves a tradeoff…

统计方法学 · 统计学 2023-02-14 Guo Yu , Daniela Witten , Jacob Bien

Conditional testing via the knockoff framework allows one to identify -- among large number of possible explanatory variables -- those that carry unique information about an outcome of interest, and also provides a false discovery rate…

统计方法学 · 统计学 2024-03-05 Benjamin B Chu , Jiaqi Gu , Zhaomeng Chen , Tim Morrison , Emmanuel Candes , Zihuai He , Chiara Sabatti

We consider the problem of variable selection in regression models. In particular, we are interested in selecting explanatory covariates linked with the response variable and we want to determine which covariates are relevant, that is which…

统计方法学 · 统计学 2019-07-09 Anne Gégout-Petit , Aurélie Gueudin-Muller , Clémence Karmann

In real-world applications of reinforcement learning, it is often challenging to obtain a state representation that is parsimonious and satisfies the Markov property without prior knowledge. Consequently, it is common practice to construct…

机器学习 · 统计学 2024-07-31 Tao Ma , Jin Zhu , Hengrui Cai , Zhengling Qi , Yunxiao Chen , Chengchun Shi , Eric B. Laber

Interpretability and stability are two important features that are desired in many contemporary big data applications arising in economics and finance. While the former is enjoyed to some extent by many existing forecasting approaches, the…

统计理论 · 数学 2018-09-14 Yingying Fan , Jinchi Lv , Mahrad Sharifvaghefi , Yoshimasa Uematsu

Variable importance plays a pivotal role in interpretable machine learning as it helps measure the impact of factors on the output of the prediction model. Model agnostic methods based on the generation of "null" features via permutation…

Controlling false discovery rate (FDR) is crucial for variable selection, multiple testing, among other signal detection problems. In literature, there is certainly no shortage of FDR control strategies when selecting individual features,…

统计方法学 · 统计学 2022-04-11 Jingyuan Liu , Ao Sun , Yuan Ke

We present a novel method for controlling the $k$-familywise error rate ($k$-FWER) in the linear regression setting using the knockoffs framework first introduced by Barber and Cand\`es. Our procedure, which we also refer to as knockoffs,…

统计方法学 · 统计学 2015-11-10 Lucas Janson , Weijie Su

The knockoff filter introduced by Barber and Cand\`es 2016 is an elegant framework for controlling the false discovery rate in variable selection. While empirical results indicate that this methodology is not too conservative, there is no…

统计理论 · 数学 2020-01-13 Jingbo Liu , Philippe Rigollet

In constructing an econometric or statistical model, we pick relevant features or variables from many candidates. A coalitional game is set up to study the selection problem where the players are the candidates and the payoff function is a…

机器学习 · 统计学 2021-10-07 Xingwei Hu