English
Related papers

Related papers: Fitting Power-laws in empirical data with estimato…

200 papers

We employ a parameter-free distribution estimation framework where estimators are random distributions and utilize the Kullback-Leibler (KL) divergence as a loss function. Wu and Vos [J. Statist. Plann. Inference 142 (2012) 1525-1536] show…

Statistics Theory · Mathematics 2015-09-21 Paul Vos , Qiang Wu

This article presents methods for estimating extreme probabilities, beyond the range of the observations. These methods are model-free and applicable to almost any sample size. They are grounded in order statistics theory and have a wide…

Applications · Statistics 2025-04-03 Joan del Castillo , Pedro Puig

In applied probability, the normal approximation is often used for the distribution of data with assumed additive structure. This tradition is based on the central limit theorem for sums of (independent) random variables. However, it is…

Probability · Mathematics 2020-10-27 Alexandra Dorofeeva , Victor Korolev , Alexander Zeifman

Designing scalable estimation algorithms is a core challenge in modern statistics. Here we introduce a framework to address this challenge based on parallel approximants, which yields estimators with provable properties that operate on the…

Methodology · Statistics 2023-08-04 Aritra Chakravorty , William S. Cleveland , Patrick J. Wolfe

We define a Maximum Likelihood (ML for short) estimator for the correlation function, {\xi}, that uses the same pair counting observables (D, R, DD, DR, RR) as the standard Landy and Szalay (1993, LS for short) estimator. The ML estimator…

Cosmology and Nongalactic Astrophysics · Physics 2013-11-27 Eric Jones Baxter , Eduardo Rozo

Random sampling is an essential tool in the processing and transmission of data. It is used to summarize data too large to store or manipulate and meet resource constraints on bandwidth or battery power. Estimators that are applied to the…

Databases · Computer Science 2015-03-19 Edith Cohen , Haim Kaplan

In a linear regression model with random design, we consider a family of candidate models from which we want to select a `good' model for prediction out-of-sample. We fit the models using block shrinkage estimators, and we focus on the…

Statistics Theory · Mathematics 2018-09-13 Hannes Leeb , Nina Senitschnig

The choice of free parameters in network models is subjective, since it depends on what topological properties are being monitored. However, we show that the Maximum Likelihood (ML) principle indicates a unique, statistically rigorous…

Disordered Systems and Neural Networks · Physics 2008-08-07 Diego Garlaschelli , Maria I. Loffredo

M-estimators are ubiquitous in machine learning and statistical learning theory. They are used both for defining prediction strategies and for evaluating their precision. In this paper, we propose the first non-asymptotic "any-time"…

Statistics Theory · Mathematics 2019-05-27 Victor-Emmanuel Brunel , Arnak S. Dalalyan , Nicolas Schreuder

Multivariate extreme value statistical analysis is concerned with observations on several variables which are thought to possess some degree of tail-dependence. In areas such as the modeling of financial and insurance risks, or as the…

Applications · Statistics 2014-12-31 Alexis Bienvenüe , Christian Y. Robert

In many settings, robust data analysis involves computational methods for uncertainty quantification and statistical inference. To design frequentist studies that leverage robust analysis methods, suitable sample sizes to achieve desired…

Methodology · Statistics 2025-12-19 Luke Hagar , Andrew J. Martin

This paper derives the nonparametric maximum likelihood estimator (NPMLE) of a distribution function from observations which are subject to both bias and censoring. The NPMLE is obtained by a simple EM algorithm which is an extension of the…

Statistics Theory · Mathematics 2007-08-22 Micha Mandel

Machine-learned interatomic potentials (MLIPs) are increasingly used to replace computationally demanding electronic-structure calculations to model matter at the atomic scale. The most commonly used model architectures are constrained to…

Chemical Physics · Physics 2026-03-30 Filippo Bigi , Paolo Pegolo , Arslan Mazitov , Jonathan Schmidt , Michele Ceriotti

Ensembles of deep neural networks are known to achieve state-of-the-art performance in uncertainty estimation and lead to accuracy improvement. In this work, we focus on a classification problem and investigate the behavior of both…

Machine Learning · Computer Science 2021-06-29 Ekaterina Lobacheva , Nadezhda Chirkova , Maxim Kodryan , Dmitry Vetrov

Maximum likelihood is the most widely used statistical estimation technique. Recent work by the authors introduced a general methodology for the construction of estimators for functionals in parametric models, and demonstrated improvements…

Methodology · Statistics 2014-09-29 Jiantao Jiao , Kartik Venkat , Yanjun Han , Tsachy Weissman

The feasibility of extrapolation of completely monotone functions can be quantified by examining the worst case scenario, whereby a pair of completely monotone functions agree on a given interval to a given relative precision, but differ as…

Complex Variables · Mathematics 2024-02-02 Henry J. Brown , Yury Grabovsky

Machine learning (ML) has emerged as a powerful tool for tackling complex regression and classification tasks, yet its success often hinges on the quality of training data. This study introduces an ML paradigm inspired by domain knowledge…

Machine Learning · Computer Science 2025-01-10 Mohsen Rashki

Some authors have recently argued that a finite-size scaling law for the text-length dependence of word-frequency distributions cannot be conceptually valid. Here we give solid quantitative evidence for the validity of such scaling law,…

Data Analysis, Statistics and Probability · Physics 2018-04-12 Alvaro Corral , Francesc Font-Clos

The power law is useful in describing count phenomena such as network degrees and word frequencies. With a single parameter, it captures the main feature that the frequencies are linear on the log-log scale. Nevertheless, there have been…

Applications · Statistics 2024-07-24 Clement Lee , Emma Eastoe , Aiden Farrell

Estimating symmetric properties of a distribution, e.g. support size, coverage, entropy, distance to uniformity, are among the most fundamental problems in algorithmic statistics. While each of these properties have been studied extensively…

Data Structures and Algorithms · Computer Science 2019-05-22 Moses Charikar , Kirankumar Shiragur , Aaron Sidford
‹ Prev 1 3 4 5 6 7 10 Next ›