English
Related papers

Related papers: Fitting heavy tailed distributions: the poweRlaw p…

200 papers

This paper introduces a new two-parameter distribution, referred to as the Shiha distribution, which provides a flexible model for skewed lifetime data with either heavy or light tails. The proposed distribution is applicable to various…

Methodology · Statistics 2026-02-04 F. A. Shiha

``When a measure becomes a target, it ceases to be a good measure'', this adage is known as {\it Goodhart's law}. In this paper, we investigate formally this law and prove that it critically depends on the tail distribution of the…

Machine Learning · Statistics 2024-10-15 El-Mahdi El-Mhamdi , Lê-Nguyên Hoang

The Lindley distribution and its numerous generalizations are widely used in statistical and engineering practice. Recently, a power transformation of Lindley distribution, called the power Lindley distribution, has been introduced by M. E.…

Statistics Theory · Mathematics 2019-07-02 Mohammed Khalleefah , Sofiya Ostrovska , Mehmet Turan

Although the fundamental probabilistic theory of extremes has been well developed, there are many practical considerations that must be addressed in application. The contribution of this thesis is four-fold. The first concerns the choice of…

Methodology · Statistics 2016-11-28 Brian Bader

Gradually Truncated Log-normal distribution - Size distribution of firms Abstract Many natural and economical phenomena are described through power law or log- normal distributions. In these cases, probability decreases very slowly with…

Statistical Mechanics · Physics 2008-12-02 Hari M. Gupta , Jose R. Campanha

Understanding the shape of a distribution of data is of interest to people in a great variety of fields, as it may affect the types of algorithms used for that data. We study one such problem in the framework of distribution property…

Machine Learning · Computer Science 2022-12-06 Maryam Aliakbarpour , Amartya Shankha Biswas , Kavya Ravichandran , Ronitt Rubinfeld

Real-world data usually present long-tailed distributions. Training on imbalanced data tends to render neural networks perform well on head classes while much worse on tail classes. The severe sparseness of training instances for the tail…

Machine Learning · Computer Science 2021-11-10 Chaozheng Wang , Shuzheng Gao , Cuiyun Gao , Pengyun Wang , Wenjie Pei , Lujia Pan , Zenglin Xu

An exact solution is presented to a model that mimics the crowding effect in financial markets which arises when groups of agents share information. We show that the size distribution of groups of agents has a power law tail with an…

Statistical Mechanics · Physics 2007-05-23 R. D'hulst , G. J. Rodgers

Identifying the statistical distribution that best fits citation data is important to allow robust and powerful quantitative analyses. Whilst previous studies have suggested that both the hooked power law and discretised lognormal…

Digital Libraries · Computer Science 2016-03-07 Mike Thelwall

In this paper we define the class of matrix Mittag-Leffler distributions and study some of its properties. We show that it can be interpreted as a particular case of an inhomogeneous phase-type distribution with random scaling factor, and…

Statistics Theory · Mathematics 2020-04-28 Hansjoerg Albrecher , Martin Bladt , Mogens Bladt

The double Pareto distribution is a heavy-tailed distribution with a power-law tail, that is generated via geometric Brownian motion with an exponentially distributed observation time. In this study, we examine a modified model wherein the…

Statistical Mechanics · Physics 2024-04-30 Ken Yamamoto , Takashi Bando , Hirokazu Yanagawa , Yoshihiro Yamazaki

This document contains the mathematical introduction to RORPack - a Python software library for robust output tracking and disturbance rejection for linear PDE systems. The RORPack library is open-source and freely available at…

Optimization and Control · Mathematics 2019-02-27 Lassi Paunonen

Motivation: Model selection is a ubiquitous challenge in statistics. For penalized models, model selection typically entails tuning hyperparameters to maximize a measure of fit or minimize out-of-sample prediction error. However, these…

Methodology · Statistics 2025-05-29 Priyam Das , Sarah Robinson , Christine B. Peterson

We introduce an \verb|R| package, called \verb|MPS|, for computing the probability density function, computing the cumulative distribution function, computing the quantile function, simulating random variables, and estimating the parameters…

Computation · Statistics 2018-09-11 Mahdi Teimouri

Ranking data represent a peculiar form of multivariate ordinal data taking values in the set of permutations. Despite the numerous methodological contributions to increase the flexibility of ranked data modeling, the application of more…

Computation · Statistics 2018-03-13 Cristina Mollica , Luca Tardella

The recursive and hierarchical structure of full rooted trees is applicable to represent statistical models in various areas, such as data compression, image processing, and machine learning. In most of these cases, the full rooted tree is…

Machine Learning · Statistics 2022-03-24 Yuta Nakahara , Shota Saito , Akira Kamatsuka , Toshiyasu Matsushima

With the rise of the "big data" phenomenon in recent years, data is coming in many different complex forms. One example of this is multi-way data that come in the form of higher-order tensors such as coloured images and movie clips.…

Methodology · Statistics 2021-06-17 Michael P. B. Gallaugher , Peter A. Tait , Paul D. McNicholas

Empirical evidence suggests that heavy-tailed degree distributions occurring in many real networks are well-approximated by power laws with exponents $\eta$ that may take values either less than and greater than two. Models based on various…

Machine Learning · Statistics 2018-07-10 Benjamin Bloem-Reddy , Adam Foster , Emile Mathieu , Yee Whye Teh

How to estimate the uncertainty of a given model is a crucial problem. Current calibration techniques treat different classes equally and thus implicitly assume that the distribution of training data is balanced, but ignore the fact that…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Jiahao Chen , Bing Su

The use of objective prior in Bayesian applications has become a common practice to analyze data without subjective information. Formal rules usually obtain these priors distributions, and the data provide the dominant information in the…

Statistics Theory · Mathematics 2020-05-18 Pedro L. Ramos , Francisco A. Rodrigues , Eduardo Ramos , Dipak K. Dey , Francisco Louzada