中文
相关论文

相关论文: Too good to be true: when overwhelming evidence fa…

200 篇论文

This paper presents Bayesian techniques for conservative claims about software reliability, particularly when evidence suggests the software's executions are not statistically independent. We formalise informal notions of "doubting" that…

软件工程 · 计算机科学 2023-10-12 Kizito Salako , Xingyu Zhao

The search for a scientific theory of consciousness should result in theories that are falsifiable. However, here we show that falsification is especially problematic for theories of consciousness. We formally describe the standard…

神经元与认知 · 定量生物学 2021-04-29 Johannes Kleiner , Erik Hoel

The plausibility of uncommon events and miracles based on testimony of such an event has been much discussed. When analyzing the probabilities involved, it has mostly been assumed that the common events can be taken as data in the…

应用统计 · 统计学 2016-03-01 V. Palonen

Modern statisticians are often presented with hundreds or thousands of hypothesis testing problems to evaluate at the same time, generated from new scientific technologies such as microarrays, medical and satellite imaging devices, or flow…

应用统计 · 统计学 2008-12-18 Bradley Efron

This paper demonstrates a methodology for examining the accuracy of uncertain inference systems (UIS), after their parameters have been optimized, and does so for several common UIS's. This methodology may be used to test the accuracy when…

人工智能 · 计算机科学 2013-04-11 Ben P. Wise

Particularly in genomics, but also in other fields, it has become commonplace to undertake highly multiple Student's $t$-tests based on relatively small sample sizes. The literature on this topic is continually expanding, but the main…

统计理论 · 数学 2010-10-11 Peter Hall , Qiying Wang

The idea of fully accepting statements when the evidence has rendered them probable enough faces a number of difficulties. We leave the interpretation of probability largely open, but attempt to suggest a contextual approach to full belief.…

人工智能 · 计算机科学 2013-02-08 Henry E. Kyburg

When assessing a software-based system, the results of Bayesian statistical inference on operational testing data can provide strong support for software reliability claims. For inference, this data (i.e. software successes and failures) is…

软件工程 · 计算机科学 2023-01-16 Kizito Salako , Xingyu Zhao

Most scientific disciplines use significance testing to draw conclusions about experimental or observational data. This classical approach provides a theoretical guarantee for controlling the number of false positives across a set of…

应用统计 · 统计学 2023-03-06 Stanley E. Lazic

Software testing is a mandatory activity in any serious software development process, as bugs are a reality in software development. This raises the question of quality: good tests are effective in finding bugs, but until a test case…

This article extends the hypotheses assessment method to the case with two competing simple hypotheses. In doing so we further clarify the benefits that hypotheses assessments can bring to classical statistical analyses. Given that…

统计方法学 · 统计学 2025-03-25 Graham N. Bornholt

While running any experiment, we often have to consider the statistical power to ensure an effective study. Statistical power or power ensures that we can observe an effect with high probability if such a true effect exists. However,…

统计方法学 · 统计学 2023-06-21 Ajinkya K Mulay , Sean Lane , Erin Hennes

This paper examines the biases and performance of several uncertain inference systems: Mycin, a variant of Mycin. and a simplified version of probability using conditional independence assumptions. We present axiomatic arguments for using…

人工智能 · 计算机科学 2013-04-12 Ben P. Wise

Statistical dependence between hypotheses poses a significant challenge to the stability of large scale multiple hypotheses testing. Ignoring it often results in an unacceptably large spread in the false positive proportion even though the…

统计方法学 · 统计学 2018-10-15 Sairam Rayaprolu , Zhiyi Chi

Adversarial examples pose a security threat to many critical systems built on neural networks. Given that deterministic robustness often comes with significantly reduced accuracy, probabilistic robustness (i.e., the probability of having…

机器学习 · 计算机科学 2024-05-27 Ruihan Zhang , Jun Sun

A popular approach to significance testing proposes to decide whether the given hypothesized statistical model is likely to be true (or false). Statistical decision theory provides a basis for this approach by requiring every significance…

统计方法学 · 统计学 2013-01-08 William Perkins , Mark Tygert , Rachel Ward

What are the criteria that a measure of statistical evidence should satisfy? It is argued that a measure of evidence should be consistent. Consistency is an asymptotic criterion: the probability that if a measure of evidence in data…

统计理论 · 数学 2011-11-22 M. Grendar

A previously proved theorem gives sufficient conditions for an estimator of the false discovery rate (FDR) to conservatively converge to the FDR with probability 1 as the number of hypothesis tests increases, even for small sample sizes. It…

基因组学 · 定量生物学 2007-05-23 David R. Bickel

We consider the problem of estimating the number of false null hypotheses among a very large number of independently tested hypotheses, focusing on the situation in which the proportion of false null hypotheses is very small. We propose a…

统计理论 · 数学 2007-06-13 Nicolai Meinshausen , John Rice

Recent research has generated hope that inference scaling, such as resampling solutions until they pass verifiers like unit tests, could allow weaker models to match stronger ones. Beyond inference, this approach also enables training…

机器学习 · 计算机科学 2026-03-27 Benedikt Stroebl , Sayash Kapoor , Arvind Narayanan
‹ 上一页 1 2 3 10 下一页 ›