English
Related papers

Related papers: Full-conformal novelty detection

200 papers

Machine learning (ML) models are powerful tools for detecting complex patterns within data, yet their "black box" nature limits their interpretability, hindering their use in critical domains like healthcare and finance. To address this…

Machine Learning · Computer Science 2025-02-17 Winston Chen , Yifan Jiang , William Stafford Noble , Yang Young Lu

With the rapid growth of crowdsourcing platforms it has become easy and relatively inexpensive to collect a dataset labeled by multiple annotators in a short time. However due to the lack of control over the quality of the annotators, some…

Machine Learning · Statistics 2016-06-17 Qianqian Xu , Jiechao Xiong , Xiaochun Cao , Yuan Yao

This article addresses a fundamental concern, first raised by Efron (2004), regarding the selection of null distributions in large-scale multiple testing. In modern data-intensive applications involving thousands or even millions of…

Methodology · Statistics 2025-09-05 Yang Tian , Zinan Zhao , Wenguang Sun

Since Benjamini and Hochberg introduced false discovery rate (FDR) in their seminal paper, this has become a very popular approach to the multiple comparisons problem. An increasingly popular topic within functional data analysis is local…

Methodology · Statistics 2019-08-16 Niels Lundtorp Olsen , Alessia Pini , Simone Vantini

Deep neural networks often predict samples with high confidence even when they come from unseen classes and should instead be flagged for expert evaluation. Current novelty detection algorithms cannot reliably identify such near OOD points…

Machine Learning · Computer Science 2022-03-14 Alexandru Ţifrea , Eric Stavarache , Fanny Yang

Novelty detection methods aim at partitioning the test units into already observed and previously unseen patterns. However, two significant issues arise: there may be considerable interest in identifying specific structures within the…

Applications · Statistics 2021-06-18 Francesco Denti , Andrea Cappozzo , Francesca Greselin

We propose the Terminating-Random Experiments (T-Rex) selector, a fast variable selection method for high-dimensional data. The T-Rex selector controls a user-defined target false discovery rate (FDR) while maximizing the number of selected…

Methodology · Statistics 2024-03-14 Jasin Machkour , Michael Muma , Daniel P. Palomar

The introduction of the false discovery rate (FDR) by Benjamini and Hochberg has spurred a great interest in developing methodologies to control the FDR in various settings. The majority of existing approaches, however, address the FDR…

Methodology · Statistics 2016-06-09 Kasra Alishahi , Ahmad Reza Ehyaei , Ali Shojaie

Online testing procedures assume that hypotheses are observed in sequence, and allow the significance thresholds for upcoming tests to depend on the test statistics observed so far. Some of the most popular online methods include alpha…

Methodology · Statistics 2022-02-11 Aaron Fisher

Multiple hypotheses testing is a core problem in statistical inference and arises in almost every scientific field. Given a sequence of null hypotheses $\mathcal{H}(n) = (H_1,..., H_n)$, Benjamini and Hochberg…

Methodology · Statistics 2015-03-05 Adel Javanmard , Andrea Montanari

To find interesting items in genome-wide association studies or next generation sequencing data, a crucial point is to design powerful false discovery rate (FDR) controlling procedures that suitably combine discrete tests (typically…

Statistics Theory · Mathematics 2017-09-18 Sebastian Döhler , Guillermo Durand , Etienne Roquain

We present false discovery rate smoothing, an empirical-Bayes method for exploiting spatial structure in large multiple-testing problems. FDR smoothing automatically finds spatially localized regions of significant test statistics. It then…

Methodology · Statistics 2016-11-15 Wesley Tansey , Oluwasanmi Koyejo , Russell A. Poldrack , James G. Scott

This work concerns controlling the false discovery rate (FDR) in networks under communication constraints. We present sample-and-forward, a flexible and communication-efficient version of the Benjamini-Hochberg (BH) procedure for multihop…

Signal Processing · Electrical Eng. & Systems 2023-05-17 Mehrdad Pournaderi , Yu Xiang

As artificial intelligence (AI) / machine learning (ML) gain widespread adoption, practitioners are increasingly seeking means to quantify and control the risk these systems incur. This challenge is especially salient when such systems have…

Machine Learning · Computer Science 2024-06-06 Drew Prinster , Samuel Stanton , Anqi Liu , Suchi Saria

We explicitly define the notions of (bona fide, approximate or asymptotic) compound p-values and e-values, which have been implicitly presented and used in the recent multiple testing literature. While it is known that the e-BH procedure…

Methodology · Statistics 2025-07-24 Nikolaos Ignatiadis , Ruodu Wang , Aaditya Ramdas

In the online multiple testing problem, p-values corresponding to different null hypotheses are observed one by one, and the decision of whether or not to reject the current hypothesis must be made immediately, after which the next p-value…

Methodology · Statistics 2017-10-03 Aaditya Ramdas , Fanny Yang , Martin J. Wainwright , Michael I. Jordan

Online multiple hypothesis testing has attracted a lot of attention in many applications, e.g., anomaly status detection and stock market price monitoring. The state-of-the-art generalized $\alpha$-investing (GAI) algorithms can control…

Methodology · Statistics 2025-08-05 Yifan Zhang , Zijian Wei , Haojie Ren , Changliang Zou

This paper proposes novel inferential procedures for discovering the network Granger causality in high-dimensional vector autoregressive models. In particular, we mainly offer two multiple testing procedures designed to control the false…

Methodology · Statistics 2024-11-14 Yoshimasa Uematsu , Takashi Yamagata

Although there is a huge literature on feature selection for the Cox model, none of the existing approaches can control the false discovery rate (FDR) unless the sample size tends to infinity. In addition, there is no formal power analysis…

Methodology · Statistics 2023-08-02 Daoji Li , Jinzhao Yu , Hui Zhao

Motivated by the genomic application of expression quantitative trait loci (eQTL) mapping, we propose a new procedure to perform simultaneous testing of multiple hypotheses using Bayes factors as input test statistics. One of the most…

Methodology · Statistics 2016-06-09 Xiaoquan Wen
‹ Prev 1 8 9 10 Next ›