中文
相关论文

相关论文: StarTrek: Combinatorial Variable Selection with Fa…

200 篇论文

We propose sequential multiple testing procedures which control the false discover rate (FDR) or the positive false discovery rate (pFDR) under arbitrary dependence between the data streams. This is accomplished by "optimizing" an upper…

统计方法学 · 统计学 2024-11-27 Michael Hankin , Jay Bartroff

Estimating conditional independence graphs from high-dimensional Gaussian data is challenging because methods must detect relevant edges while rigorously controlling statistical errors. We propose a Bayesian framework based on a prior…

统计方法学 · 统计学 2026-04-21 Roland B. Sogan , Tabea Rebafka , Fanny Villers

We propose the group knockoff filter, a method for false discovery rate control in a linear regression setting where the features are grouped, and we would like to select a set of relevant groups which have a nonzero effect on the response.…

统计方法学 · 统计学 2016-02-12 Ran Dai , Rina Foygel Barber

Network filtering is an important form of dimension reduction to isolate the core constituents of large and interconnected complex systems. We introduce a new technique to filter large dimensional networks arising out of dynamical behavior…

机器学习 · 统计学 2021-01-25 Arnab Chakrabarti , Anindya S. Chakrabarti

As the volume and complexity of data continue to expand across various scientific disciplines, the need for robust methods to account for the multiplicity of comparisons has grown widespread. A popular measure of type 1 error rate in…

统计方法学 · 统计学 2024-11-19 Jianliang He , Bowen Gang , Luella Fu

This paper proposes a model-free and data-adaptive feature screening method for ultra-high dimensional datasets. The proposed method is based on the projection correlation which measures the dependence between two random vectors. This…

统计方法学 · 统计学 2021-02-16 Wanjun Liu , Yuan Ke , Jingyuan Liu , Runze Li

In the last few years, compression of deep neural networks has become an important strand of machine learning and computer vision research. Deep models require sizeable computational complexity and storage, when used for instance for Human…

计算机视觉与模式识别 · 计算机科学 2020-11-10 Ayush Srivastava , Oshin Dutta , Prathosh AP , Sumeet Agarwal , Jigyasa Gupta

Given a network, the critical node detection problem finds a subset of nodes whose removal disrupts the network connectivity. Since many real-world systems are naturally modeled as graphs, assessing the vulnerability of the network is…

离散数学 · 计算机科学 2025-12-02 Tuguldur Bayarsaikhan , Altannar Chinchuluun , Ashwin Arulselvan , Panos Pardalos

This paper studies parametric bootstrap methods for network data, with the goal of quantifying the uncertainty of network statistics of interest. While existing network resampling methods primarily focus on count statistics under…

统计方法学 · 统计学 2026-05-29 Zhixuan Shao , Can M. Le

We are given a set of $n$ points that might be uniformly distributed in the unit square $[0,1]^2$. We wish to test whether the set, although mostly consisting of uniformly scattered points, also contains a small fraction of points sampled…

统计理论 · 数学 2007-06-13 Ery Arias-Castro , David L. Donoho , Xiaoming Huo

The goal of Feature Selection - comprising filter, wrapper, and embedded approaches - is to find the optimal feature subset for designated downstream tasks. Nevertheless, current feature selection methods are limited by: 1) the selection…

机器学习 · 计算机科学 2023-09-18 Meng Xiao , Dongjie Wang , Min Wu , Pengfei Wang , Yuanchun Zhou , Yanjie Fu

One challenge in exploratory association studies using observational data is that the associations between the predictors and the outcome are potentially weak and rare, and the candidate predictors have complex correlation structures. False…

统计方法学 · 统计学 2025-01-30 Runqiu Wang , Ran Dai , Hongying Dai , Evan French , Cheng Zheng

Motivation: Target-decoy search (TDS) is currently the most popular strategy for estimating and controlling the false discovery rate (FDR) of peptide identifications in mass spectrometry-based shotgun proteomics. While this strategy is very…

应用统计 · 统计学 2015-01-06 Kun He , Yan Fu , Wen-Feng Zeng , Lan Luo , Hao Chi , Chao Liu , Lai-Yun Qing , Rui-Xiang Sun , Si-Min He

Simultaneously finding multiple influential variables and controlling the false discovery rate (FDR) for linear regression models is a fundamental problem. We here propose the Gaussian Mirror (GM) method, which creates for each predictor…

统计方法学 · 统计学 2021-03-22 Xin Xing , Zhigen Zhao , Jun S. Liu

This work addresses the critical challenge of optimal filter selection for a novel trace gas measurement device. This device uses photonic crystal filters to retrieve trace gas concentrations affected by photon and read noise. The filter…

神经与进化计算 · 计算机科学 2025-11-10 Kirill Antonov , Marijn Siemons , Niki van Stein , Thomas H. W. Bäck , Ralf Kohlhaas , Anna V. Kononova

Ground-based optical surveys such as PanSTARRS, DES, and LSST, will produce large catalogs to limiting magnitudes of r > 24. Star-galaxy separation poses a major challenge to such surveys because galaxies---even very compact…

天体物理仪器与方法 · 物理学 2015-06-05 Ross Fadely , David W. Hogg , Beth Willman

False discovery rate (FDR) procedures provide misleading inference when testing multiple null hypotheses with heterogeneous multinomial data. For example, in the motivating study the goal is to identify species of bacteria near the roots of…

统计方法学 · 统计学 2015-11-05 Joshua Habiger , David Watts , Michael Anderson

Unsupervised feature selection is an important method to reduce dimensions of high dimensional data without labels, which is benefit to avoid ``curse of dimensionality'' and improve the performance of subsequent machine learning tasks, like…

机器学习 · 计算机科学 2020-12-29 Yanyong Huang , Zongxin Shen , Fuxu Cai , Tianrui Li , Fengmao Lv

Pedestrian trajectory prediction is crucial for many important applications. This problem is a great challenge because of complicated interactions among pedestrians. Previous methods model only the pairwise interactions between pedestrians,…

计算机视觉与模式识别 · 计算机科学 2020-01-14 Yanliang Zhu , Deheng Qian , Dongchun Ren , Huaxia Xia

Two-sample hypothesis testing for network comparison presents many significant challenges, including: leveraging repeated network observations and known node registration, but without requiring them to operate; relaxing strong structural…

统计方法学 · 统计学 2024-02-05 Meijia Shao , Dong Xia , Yuan Zhang , Qiong Wu , Shuo Chen