中文
相关论文

相关论文: Divergence vs. Decision P-values: A Distinction Wo…

200 篇论文

Attempts to replicate probabilistic reasoning in expert systems have typically overlooked a critical ingredient of that process. Probabilistic analysis typically requires extensive judgments regarding interdependencies among hypotheses and…

人工智能 · 计算机科学 2013-04-15 Marvin S. Cohen

Measurement devices always add noise to the signal of interest and it is necessary to evaluate the variance of the results. This article focuses on stationary random processes whose Power Spectrum Density is a power law of frequency. For…

数据分析、统计与概率 · 物理学 2013-05-20 Benjamin Lenoir

Statistical significance of both the original and the replication study is a commonly used criterion to assess replication attempts, also known as the two-trials rule in drug development. However, replication studies are sometimes conducted…

应用统计 · 统计学 2024-05-31 Leonhard Held , Samuel Pawel , Charlotte Micheloud

Our interest is whether two binomial parameters differ, which parameter is larger, and by how much. This apparently simple problem was addressed by Fisher in the 1930's, and has been the subject of many review papers since then. Yet there…

统计方法学 · 统计学 2021-04-20 Michael P. Fay , Sally A. Hunsberger

Recent works have shown that machine learning models improve at a predictable rate with the total amount of training data, leading to scaling laws that describe the relationship between error and dataset size. These scaling laws can help…

机器学习 · 计算机科学 2024-06-03 Ian Covert , Wenlong Ji , Tatsunori Hashimoto , James Zou

A definition for the statistical significance by constructing a correlation between the normal distribution integral probability and the p-value observed in an experiment is proposed, which is suitable for both counting experiment and…

数据分析、统计与概率 · 物理学 2007-05-23 Yongsheng Zhu

Estimating the difficulty of a dataset typically involves comparing state-of-the-art models to humans; the bigger the performance gap, the harder the dataset is said to be. However, this comparison provides little understanding of how…

计算与语言 · 计算机科学 2025-04-29 Kawin Ethayarajh , Yejin Choi , Swabha Swayamdipta

Software development processes are subject to variations in time and space, variations that can originate from learning effects, differences in application domains, or a number of other causes. Identifying and analyzing such differences is…

软件工程 · 计算机科学 2014-01-21 Martín Soto , Jürgen Münch

Examining residuals such as Pearson and deviance residuals, is a standard tool for assessing normal regression. However, for discrete response, these residuals cluster on lines corresponding to distinct response values. Their distributions…

统计方法学 · 统计学 2020-07-07 Cindy Feng , Alireza Sadeghpour , Longhai Li

Null hypothesis statistical significance testing (NHST) is the dominant approach for evaluating results from randomized controlled trials. Whereas NHST comes with long-run error rate guarantees, its main inferential tool -- the $p$-value --…

统计方法学 · 统计学 2022-06-10 František Bartoš , Samuel Pawel , Eric-Jan Wagenmakers

Observations on the past provide some hints about what will happen in the future, and this can be quantified using information theory. The ``predictive information'' defined in this way has connections to measures of complexity that have…

统计力学 · 物理学 2007-05-23 William Bialek , Naftali Tishby

The purpose of the paper is to provide a new way of seeing the p-value in terms of a fuzzy membership function. According to the ASAs statement, we aim at removing the arbitrary choice of the significance level and at demonstrating that the…

其他统计学 · 统计学 2025-08-12 Piero Quatto

This article develops $p$-values for evaluating means of normal populations that make use of indirect or prior information. A $p$-value of this type is based on a biased test statistic that is optimal on average with respect to a…

统计方法学 · 统计学 2019-12-12 Peter D. Hoff

The general relationship between an arbitrary frequency distribution and the expectation value of the frequency distributions of its samples is esablished. A set of combinations of expectation values whose value does not in general depend…

数据分析、统计与概率 · 物理学 2012-10-05 Paolo Rossi

For a given testing problem, let $U_1,...,U_n$ be individually valid and conditionally on the data i.i.d.\ P-variables (often called P-values). For example, the data could come in groups, and each $U_i$ could be based on subsampling just…

统计方法学 · 统计学 2011-08-22 Lutz Mattner

This paper is an attempt to set a justification for making use of some dicrepancy indexes, starting from the classical Maximum Likelihood definition, and adapting the corresponding basic principle of inference to situations where…

统计理论 · 数学 2021-02-24 Michel Broniatowski

In this paper we use e-values in the context of multiple hypothesis testing assuming that the base tests produce independent, or sequential, e-values. Our simulation and empirical studies and theoretical considerations suggest that, under…

统计方法学 · 统计学 2024-08-14 Vladimir Vovk , Ruodu Wang

Model averaging is a useful and robust method for dealing with model uncertainty in statistical analysis. Often, it is useful to consider data subset selection at the same time, in which model selection criteria are used to compare models…

统计方法学 · 统计学 2023-10-26 Ethan T. Neil , Jacob W. Sitison

Calibration$\unicode{x2014}$the problem of ensuring that predicted probabilities align with observed class frequencies$\unicode{x2014}$is a basic desideratum for reliable prediction with machine learning systems. Calibration error is…

机器学习 · 统计学 2026-03-02 Eugène Berta , Sacha Braun , David Holzmüller , Francis Bach , Michael I. Jordan

It is argued that all model based approaches to the selection of covariates in linear regression have failed. This applies to frequentist approaches based on P-values and to Bayesian approaches although for different reasons. In the first…

统计方法学 · 统计学 2022-02-23 Laurie Davies
‹ 上一页 1 8 9 10 下一页 ›