中文
相关论文

相关论文: BOOST: Power-Optimal Strong-FWER Testing for Block…

200 篇论文

Due to the vast testing space, the increasing demand for effective and efficient testing of deep neural networks (DNNs) has led to the development of various DNN test case prioritization techniques. However, the fact that DNNs can deliver…

软件工程 · 计算机科学 2024-09-17 Jialuo Chen , Jingyi Wang , Xiyue Zhang , Youcheng Sun , Marta Kwiatkowska , Jiming Chen , Peng Cheng

We propose a new adaptive empirical Bayes framework, the Bag-Of-Null-Statistics (BONuS) procedure, for multiple testing where each hypothesis testing problem is itself multivariate or nonparametric. BONuS is an adaptive and interactive…

统计方法学 · 统计学 2021-07-05 Chiao-Yu Yang , Lihua Lei , Nhat Ho , Will Fithian

We propose a novel technique to boost the power of testing a high-dimensional vector $H:\btheta=0$ against sparse alternatives where the null hypothesis is violated only by a couple of components. Existing tests based on quadratic forms…

统计方法学 · 统计学 2014-08-19 Jianqing Fan , Yuan Liao , Jiawei Yao

In this paper we introduce a novel procedure for improving multiple testing procedures (MTPs) under scenarios when the null hypothesis $p$-values tend to be stochastically larger than standard uniform (referred to as 'inflated'). An…

统计方法学 · 统计学 2025-08-29 Jules L. Ellis , Jakub Pecanka , Jelle Goeman

The problem of adversarial robustness has been studied extensively for neural networks. However, for boosted decision trees and decision stumps there are almost no results, even though they are widely used in practice (e.g. XGBoost) due to…

机器学习 · 计算机科学 2019-11-01 Maksym Andriushchenko , Matthias Hein

Designs for Order-of-Addition (OofA) experiments have received growing attention due to their impact on responses based on the sequence of component addition. In certain cases, these experiments involve heterogeneous groups of units, which…

统计方法学 · 统计学 2026-02-04 Chang-Yun Lin

Consider the multiple testing problem of testing null hypotheses $H_1,...,H_s$. A classical approach to dealing with the multiplicity problem is to restrict attention to procedures that control the familywise error rate ($\mathit{FWER}$),…

统计理论 · 数学 2007-06-13 Joseph P. Romano , Azeem M. Shaikh

Multiple hypothesis testing problems arise naturally in science. In this paper, we introduce the new Fast Closed Testing (FACT) method for multiple testing, controlling the family-wise error rate. This error rate is state of the art in many…

统计方法学 · 统计学 2020-01-22 Edgar Dobriban

This paper establishes a precise high-dimensional asymptotic theory for boosting on separable data, taking statistical and computational perspectives. We consider a high-dimensional setting where the number of features (weak learners) $p$…

统计理论 · 数学 2022-11-21 Tengyuan Liang , Pragya Sur

Blocking is a mechanism to improve the efficiency of Entity Resolution (ER) which aims to quickly prune out all non-matching record pairs. However, depending on the distributions of entity cluster sizes, existing techniques can be either…

数据库 · 计算机科学 2021-03-17 Sainyam Galhotra , Donatella Firmani , Barna Saha , Divesh Srivastava

We study a structured permutation scheme for two-sample testing that restricts permutations to single cross-swaps between block-selected representatives. Our analysis yields three main results. First, we provide an exact validity…

机器学习 · 统计学 2025-12-02 Jungwoo Ho

We present a data-driven modeling and control framework for physics-based building emulators. Our approach consists of: (a) Offline training of differentiable surrogate models that accelerate model evaluations, provide cost-effective…

系统与控制 · 电气工程与系统科学 2024-04-03 Saman Mostafavi , Chihyeon Song , Aayushman Sharma , Raman Goyal , Alejandro Brito

We analyze control of the familywise error rate (FWER) in a multiple testing scenario with a great many null hypotheses about the distribution of a high-dimensional random variable among which only a very small fraction are false, or…

统计方法学 · 统计学 2015-09-15 Kamel Lahouel , Donald Geman , Laurent Younes

Consider the problem of testing $s$ hypotheses simultaneously. The usual approach restricts attention to procedures that control the probability of even one false rejection, the familywise error rate (FWER). If $s$ is large, one might be…

统计理论 · 数学 2007-11-06 Joseph P. Romano , Michael Wolf

In many applications of multiple hypothesis testing where more than one false rejection can be tolerated, procedures controlling error rates measuring at least $k$ false rejections, instead of at least one, for some fixed $k\ge 1$ can…

统计理论 · 数学 2008-12-18 Sanat K. Sarkar

Recent deep clustering models have produced impressive clustering performance. However, a common issue with existing methods is the disparity between global and local feature structures. While local structures typically show strong…

计算机视觉与模式识别 · 计算机科学 2025-11-27 Hanyang Li , Yuheng Jia , Hui Liu , Junhui Hou

ProBoost, a new boosting algorithm for probabilistic classifiers, is proposed in this work. This algorithm uses the epistemic uncertainty of each training sample to determine the most challenging/uncertain ones; the relevance of these…

This article considers the problem of multiple hypothesis testing using $t$-tests. The observed data are assumed to be independently generated conditional on an underlying and unknown two-state hidden model. We propose an asymptotically…

统计理论 · 数学 2011-02-22 Hongyuan Cao , Michael R. Kosorok

Pattern recognition applications often suffer from skewed data distributions between classes, which may vary during operations w.r.t. the design data. Two-class classification systems designed using skewed data tend to recognize the…

机器学习 · 计算机科学 2019-12-02 Roghayeh Soleymani , Eric Granger , Giorgio Fumera

Experimental evaluations of public policies often randomize a new intervention within many sites or blocks. After a report of an overall result -- statistically significant or not -- the natural question from a policy maker is: \emph{where}…

统计方法学 · 统计学 2026-03-02 Jake Bowers , David Kim , Nuole Chen