English
Related papers

Related papers: Improved Component Predictions of Batting Measures

200 papers

The Dirichlet distribution, also known as multivariate beta, is the most used to analyse frequencies or proportions data. Maximum likelihood is widespread for estimation of Dirichlet's parameters. However, for small sample sizes, the…

Methodology · Statistics 2021-03-04 Vincenzo Gioia , Euloge Clovis Kenne Pagui

The Pythagorean formula is one of the most popular ways to measure the true ability of a team. It is very easy to use, estimating a team's winning percentage from the runs they score and allow. This data is readily available on standings…

History and Overview · Mathematics 2014-06-04 Steven J. Miller , Taylor Corcoran , Jennifer Gossels , Victor Luo , Jaclyn Porfilio

When assessing the causal effect of a binary exposure using observational data, confounder imbalance across exposure arms must be addressed. Matching methods, including propensity score-based matching, can be used to deconfound the causal…

Methodology · Statistics 2024-10-01 Ernesto Ulloa-Pérez , Marco Carone , Alex Luedtke

A mixture with varying concentrations is a modification of a finite mixture model in which the mixing probabilities (concentrations of mixture components) may be different for different observations. In the paper, we assume that the…

Probability · Mathematics 2015-03-19 Alexey Doronin , Rostyslav Maiboroda

This paper proposes new estimators for the propensity score that aim to maximize the covariate distribution balance among different treatment groups. Heuristically, our proposed procedure attempts to estimate a propensity score model by…

Econometrics · Economics 2020-04-07 Pedro H. C. Sant'Anna , Xiaojun Song , Qi Xu

There are various measures of predictive uncertainty in the literature, but their relationships to each other remain unclear. This paper uses a decomposition of statistical pointwise risk into components, associated with different sources…

Machine Learning · Statistics 2025-02-18 Nikita Kotelevskii , Vladimir Kondratyev , Martin Takáč , Éric Moulines , Maxim Panov

The Price equation partitions the change in the expected value of a population measure. The first component describes the partial change caused by altered frequencies. The second component describes the partial change caused by altered…

Populations and Evolution · Quantitative Biology 2020-03-23 Steven A. Frank , William Godsoe

Betting markets are gaining in popularity. Mean beliefs generally differ from prices in prediction markets. Logarithmic utility is employed to study the risk and return adjustments to prices. Some consequences are described. A modified…

Portfolio Management · Quantitative Finance 2024-12-19 Bernhard K Meister

Several performance measures can be used for evaluating classification results: accuracy, F-measure, and many others. Can we say that some of them are better than others, or, ideally, choose one measure that is best in all situations? To…

Machine Learning · Computer Science 2022-01-25 Martijn Gösgens , Anton Zhiyanov , Alexey Tikhonov , Liudmila Prokhorenkova

In the high-stakes world of baseball, every nuance of a pitcher's mechanics holds the key to maximizing performance and minimizing runs. Traditional analysis methods often rely on pre-recorded offline numerical data, hindering their…

Computer Vision and Pattern Recognition · Computer Science 2024-05-14 Jerrin Bright , Bavesh Balaji , Yuhao Chen , David A Clausi , John S Zelek

Cluster randomization trials commonly employ multiple endpoints. When a single summary of treatment effects across endpoints is of primary interest, global hypothesis testing/effect estimation methods represent a common analysis strategy.…

Methodology · Statistics 2025-05-19 E. Davies Smith , V. Jairath , G. Zou

The record statistics of complex random states are analytically calculated, and shown that the probability of a record intensity is a Bernoulli process. The correlation due to normalization leads to a probability distribution of the records…

Statistical Mechanics · Physics 2015-06-05 Shashi C. L. Srivastava , Arul Lakshminarayan , Sudhir R. Jain

Gibbs-type random probability measures and the exchangeable random partitions they induce represent an important framework both from a theoretical and applied point of view. In the present paper, motivated by species sampling problems, we…

Probability · Mathematics 2013-09-06 Stefano Favaro , Antonio Lijoi , Igor Prünster

Forecasting the popularity of new songs has become a standard practice in the music industry and provides a comparative advantage for those that do it well. Considerable efforts were put into machine learning prediction models for that…

Physics and Society · Physics 2022-11-29 Niklas Reisz , Vito D. P. Servedio , Stefan Thurner

Shot charts in basketball analytics provide an indispensable tool for evaluating players' shooting performance by visually representing the distribution of field goal attempts across different court locations. However, conventional methods…

Methodology · Statistics 2025-05-16 Luca Scrucca , Dimitris Karlis

Power-meter measurements are used to study a model that accounts for the use of power by a cyclist. The focus is on relations between rates of change of model quantities, such as power and speed, both in the context of partial derivatives,…

Popular Physics · Physics 2020-12-25 Tomasz Danek , Michael A. Slawinski , Theodore Stanoev

Principal component analysis (PCA) is commonly used in genetics to infer and visualize population structure and admixture between populations. PCA is often interpreted in a way similar to inferred admixture proportions, where it is assumed…

Methodology · Statistics 2023-02-10 Jan van Waaij , Song Li , Genís Garcia-Erill , Anders Albrechtsen , Carsten Wiuf

In this paper we give a brief review of semiparametric theory, using as a running example the common problem of estimating an average causal effect. Semiparametric models allow at least part of the data-generating process to be unspecified…

Methodology · Statistics 2017-09-20 Edward H. Kennedy

Compression of integer sets and sequences has been extensively studied for settings where elements follow a uniform probability distribution. In addition, methods exist that exploit clustering of elements in order to achieve higher…

Information Theory · Computer Science 2014-02-11 N. Jesper Larsson

This paper describes measures for evaluating the three determinants of how well a probabilistic classifier performs on a given test set. These determinants are the appropriateness, for the test set, of the results of (1) feature selection,…

cmp-lg · Computer Science 2008-02-03 Rebecca Bruce , Janyce Wiebe , Ted Pedersen
‹ Prev 1 8 9 10 Next ›