中文
相关论文

相关论文: Stab-GKnock: Controlled variable selection for par…

200 篇论文

We consider problems where many, somewhat redundant, hypotheses are tested and we are interested in reporting the most precise rejections, with false discovery rate (FDR) control. This is the case, for example, when researchers are…

统计方法学 · 统计学 2024-04-23 Paula Gablenz , Chiara Sabatti

The generalized linear models (GLM) have been widely used in practice to model non-Gaussian response variables. When the number of explanatory features is relatively large, scientific researchers are of interest to perform controlled…

统计方法学 · 统计学 2020-07-03 Chenguang Dai , Buyu Lin , Xin Xing , Jun S. Liu

The recent paper Cand\`es et al. (2018) introduced model-X knockoffs, a method for variable selection that provably and non-asymptotically controls the false discovery rate with no restrictions or assumptions on the dimensionality of the…

统计方法学 · 统计学 2020-06-16 Dongming Huang , Lucas Janson

An important problem in machine learning and statistics is to identify features that causally affect the outcome. This is often impossible to do from purely observational data, and a natural relaxation is to identify features that are…

机器学习 · 统计学 2019-05-30 Jaime Roquero Gimenez , Amirata Ghorbani , James Zou

This paper studies the estimation of high dimensional Gaussian graphical model (GGM). Typically, the existing methods depend on regularization techniques. As a result, it is necessary to choose the regularized parameter. However, the…

统计方法学 · 统计学 2013-06-06 Weidong Liu

Selecting important features that have substantial effects on the response with provable type-I error rate control is a fundamental concern in statistics, with wide-ranging practical applications. Existing knockoff filters, although shown…

统计方法学 · 统计学 2024-02-29 Jiaqi Gu , Zihuai He

Identifying which variables do influence a response while controlling false positives pervades statistics and data science. In this paper, we consider a scenario in which we only have access to summary statistics, such as the values of…

统计方法学 · 统计学 2024-02-21 Zhaomeng Chen , Zihuai He , Benjamin B. Chu , Jiaqi Gu , Tim Morrison , Chiara Sabatti , Emmanuel Candès

Selecting relevant features associated with a given response variable is an important issue in many scientific fields. Quantifying quality and uncertainty of a selection result via false discovery rate (FDR) control has been of recent…

统计方法学 · 统计学 2020-12-17 Chenguang Dai , Buyu Lin , Xin Xing , Jun S. Liu

Quantifying how genomic features influence different parts of an outcome distribution requires statistical tools that go beyond mean regression, especially in ultrahigh-dimensional settings. Motivated by the study of LINE-1 activity in…

统计方法学 · 统计学 2025-11-27 Sang Kyu Lee , Tongwu Zhang , Hyokyoung G. Hong , Haolei Weng

We make some initial attempt to establish the theoretical and methodological foundation for the model-X knockoffs inference for time series data. We suggest the method of time series knockoffs inference (TSKI) by exploiting the ideas of…

统计方法学 · 统计学 2025-03-03 Chien-Ming Chi , Yingying Fan , Ching-Kang Ing , Jinchi Lv

A core strength of knockoff methods is their virtually limitless customizability, allowing an analyst to exploit machine learning algorithms and domain knowledge without threatening the method's robust finite-sample false discovery rate…

统计理论 · 数学 2021-07-15 Xiao Li , William Fithian

This paper studies the distributed conditional feature screening for massive data with ultrahigh-dimensional features. Specifically, three distributed partial correlation feature screening methods (SAPS, ACPS and JDPS methods) are firstly…

统计方法学 · 统计学 2024-03-12 Naiwen Pang , Xiaochao Xia

Controlling the false discovery rate (FDR) in high-dimensional variable selection requires balancing rigorous error control with statistical power. Existing methods with provable guarantees are often overly conservative, creating a…

统计方法学 · 统计学 2026-02-06 Arnau Vilella , Jasin Machkour , Michael Muma , Daniel P. Palomar

In high dimensional variable selection problems, statisticians often seek to design multiple testing procedures that control the False Discovery Rate (FDR), while concurrently identifying a greater number of relevant variables. Model-X…

统计理论 · 数学 2023-07-25 Taejoo Ahn , Licong Lin , Song Mei

The $\gamma$-FDP and $k$-FWER multiple testing error metrics, which are tail probabilities of the respective error statistics, have become popular recently as less-stringent alternatives to the FDR and FWER. We propose general and flexible…

统计方法学 · 统计学 2016-12-20 Jay Bartroff

In real-world applications of reinforcement learning, it is often challenging to obtain a state representation that is parsimonious and satisfies the Markov property without prior knowledge. Consequently, it is common practice to construct…

机器学习 · 统计学 2024-07-31 Tao Ma , Jin Zhu , Hengrui Cai , Zhengling Qi , Yunxiao Chen , Chengchun Shi , Eric B. Laber

Effectively controlling the false discovery rate (FDR) in high-dimensional variable selection is a fundamental statistical problem that has garnered significant research interest. In this paper, we propose a novel, user-friendly, and…

统计方法学 · 统计学 2026-04-28 Yujia Wu , Panxu Yuan , Binyan Jiang

Balancing false discovery rate (FDR) control with high statistical power remains a central challenge in high-dimensional variable selection. While several FDR-controlling methods have been proposed, many degrade the original data -- by…

统计方法学 · 统计学 2025-07-16 Changhu Wang , Ziheng Zhang , Jingyi Jessica Li

Interpretability and stability are two important features that are desired in many contemporary big data applications arising in economics and finance. While the former is enjoyed to some extent by many existing forecasting approaches, the…

统计理论 · 数学 2018-09-14 Yingying Fan , Jinchi Lv , Mahrad Sharifvaghefi , Yoshimasa Uematsu

Identifying truly predictive covariates while strictly controlling false discoveries remains a fundamental challenge in nonlinear, highly correlated, and low signal-to-noise regimes, where deep learning based feature selection methods are…

机器学习 · 计算机科学 2026-02-03 Bob Junyi Zou , Lu Tian