English
Related papers

Related papers: A Generalized Publication Bias Model

200 papers

A central challenge in statistical inference is the presence of confounding variables that may distort observed associations between treatment and outcome. Conventional "causal" methods, grounded in assumptions such as ignorability, exclude…

Methodology · Statistics 2025-09-09 Ellis Scharfenaker , Duncan K. Foley

This study introduces an approach to estimate the uncertainty in bibliometric indicator values that is caused by data errors. This approach utilizes Bayesian regression models, estimated from empirical data samples, which are used to…

Digital Libraries · Computer Science 2024-12-11 Paul Donner

The recently proposed fractional scoring scheme is used to attribute publications to percentile rank classes. It is shown that in this way uncertainties and ambiguities in the evaluation of percentile ranks do not occur. Using the…

Other Statistics · Statistics 2012-05-17 Michael Schreiber

The generalized negative binomial distribution (GNB) is a new flexible family of discrete distributions that are mixed Poisson laws with the mixing generalized gamma (GG) distributions. This family of discrete distributions is very wide and…

Methodology · Statistics 2018-09-18 Andrey K. Gorshenin , Victor Yu. Korolev

The proportion of false null hypotheses is a very important quantity in statistical modelling and inference based on the two-component mixture model and its extensions, and in control and estimation of the false discovery rate and false…

Probability · Mathematics 2020-03-09 Xiongzhi Chen

Network data is prevalent in many contemporary big data applications in which a common interest is to unveil important latent links between different pairs of nodes. Yet a simple fundamental question of how to precisely quantify the…

Methodology · Statistics 2021-08-31 Jianqing Fan , Yingying Fan , Xiao Han , Jinchi Lv

By extrapolating the explicit formula of the zero-bias distribution occurring in the context of Stein's method, we construct characterization identities for a large class of absolutely continuous univariate distributions. Instead of trying…

Statistics Theory · Mathematics 2021-02-26 Steffen Betsch , Bruno Ebner

Machine Learning techniques have become pervasive across a range of different applications, and are now widely used in areas as disparate as recidivism prediction, consumer credit-risk analysis and insurance pricing. The prevalence of…

Machine Learning · Computer Science 2020-01-14 Michael Varley , Vaishak Belle

Shannon defined the mutual information between two variables. We illustrate why the true mutual information between a variable and the predictions made by a prediction algorithm is not a suitable measure of prediction quality, but the…

Statistics Theory · Mathematics 2025-04-11 Roger Sewell

Finite Sample Smeariness (FSS) has been recently discovered. It means that the distribution of sample Fr\'echet means of underlying rather unsuspicious random variables can behave as if it were smeary for quite large regimes of finite…

Statistics Theory · Mathematics 2021-03-02 Benjamin Eltzner , Shayan Hundrieser , Stephan F. Huckemann

The ratio of Bayesian evidences is a popular tool in cosmology to compare different models. There are however several issues with this method: Bayes' ratio depends on the prior even in the limit of non-informative priors, and Jeffrey's…

Cosmology and Nongalactic Astrophysics · Physics 2024-12-16 Luca Amendola , Vrund Patel , Ziad Sakr , Elena Sellentin , Kevin Wolz

Few-Shot Classification(FSC) aims to generalize from base classes to novel classes given very limited labeled samples, which is an important step on the path toward human-like machine learning. State-of-the-art solutions involve learning to…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Xiongkun Linghu , Yan Bai , Yihang Lou , Shengsen Wu , Jinze Li , Jianzhong He , Tao Bai

This letter studies a distribution-free, finite-sample data perturbation (DP) method, the Residual-Permuted Sums (RPS), which is an alternative of the Sign-Perturbed Sums (SPS) algorithm, to construct confidence regions. While SPS assumes…

Systems and Control · Electrical Eng. & Systems 2024-06-11 Szabolcs Szentpéteri , Balázs Csanád Csáji

In a recent paper (Efron (2004)), Efron pointed out that an important issue in large-scale multiple hypothesis testing is that the null distribution may be unknown and need to be estimated. Consider a Gaussian mixture model, where the null…

Statistics Theory · Mathematics 2009-11-20 Jiashun Jin , Jie Peng , Pei Wang

We study the problem, introduced by Qiao and Valiant, of learning from untrusted batches. Here, we assume $m$ users, all of whom have samples from some underlying distribution $p$ over $1, \ldots, n$. Each user sends a batch of $k$ i.i.d.…

Data Structures and Algorithms · Computer Science 2019-11-07 Sitan Chen , Jerry Li , Ankur Moitra

Modern statistical software and machine learning libraries are enabling semi-automated statistical inference. Within this context, it appears easier and easier to try and fit many models to the data at hand, reversing thereby the Fisherian…

Methodology · Statistics 2020-09-28 Pierre-Alexandre Mattei

In the past two decades, psychological science has experienced an unprecedented replicability crisis which uncovered several issues. Among others, statistical inference is too often viewed as an isolated procedure limited to the analysis of…

In this article, we investigate posterior convergence in nonparametric regression models where the unknown regression function is modeled by some appropriate stochastic process. In this regard, we consider two setups. The first setup is…

Statistics Theory · Mathematics 2020-05-04 Debashis Chatterjee , Sourabh Bhattacharya

A common approach to statistical learning with big-data is to randomly split it among $m$ machines and learn the parameter of interest by averaging the $m$ individual estimates. In this paper, focusing on empirical risk minimization, or…

Machine Learning · Statistics 2016-06-14 Jonathan Rosenblatt , Boaz Nadler

The reliability of the results of network meta-analysis (NMA) lies in the plausibility of key assumption of transitivity. This assumption implies that the effect modifiers' distribution is similar across treatment comparisons. Transitivity…

‹ Prev 1 8 9 10 Next ›