English
Related papers

Related papers: Bayesian Properties of Normalized Maximum Likeliho…

200 papers

We introduce entropic strict minimum message length (SMML), a risk-sensitive generalization of strict minimum message length coding. The proposed criterion replaces expected two-part codelength under the prior predictive distribution with…

Statistics Theory · Mathematics 2026-05-20 Enes Makalic , Daniel F. Schmidt

Maximum likelihood estimation (MLE) is a fundamental computational problem in statistics. In this paper, MLE for statistical models with discrete data is studied from an algebraic statistics viewpoint. A reformulation of the MLE problem in…

Statistics Theory · Mathematics 2014-05-27 Jose Israel Rodriguez

Maximum likelihood estimation is a common method of estimating the parameters of the probability distribution from a given sample. This paper aims to introduce the maximum likelihood estimation in the framework of sublinear expectation. We…

Probability · Mathematics 2023-01-16 Xinpeng Li , Yue Liu , Jiaquan Lu

We study the problems of data compression, gambling and prediction of a sequence $x^n=x_1x_2...x_n$ from an alphabet ${\cal X}$, in terms of regret and expected regret (redundancy) with respect to various smooth families of probability…

Information Theory · Computer Science 2025-07-25 Jun'ichi Takeuchi , Andrew R. Barron

Modern challenges of robustness, fairness, and decision-making in machine learning have led to the formulation of multi-distribution learning (MDL) frameworks in which a predictor is optimized across multiple distributions. We study the…

Machine Learning · Computer Science 2024-12-19 Rajeev Verma , Volker Fischer , Eric Nalisnick

As an alternative to variable selection or shrinkage in high dimensional regression, we propose to randomly compress the predictors prior to analysis. This dramatically reduces storage and computational bottlenecks, performing well when the…

Machine Learning · Statistics 2013-03-26 Rajarshi Guhaniyogi , David B. Dunson

We study the properties of the Minimum Description Length principle for sequence prediction, considering a two-part MDL estimator which is chosen from a countable class of models. This applies in particular to the important case of…

Machine Learning · Computer Science 2011-11-09 Jan Poland , Marcus Hutter

In this work, we derive some novel properties of the bimodal normal distribution. Some of its mathematical properties are examined. We provide a formal proof for the bimodality and assess identifiability. We then discuss the maximum…

Statistics Theory · Mathematics 2021-06-02 Roberto Vila , Helton Saulo , Jamer Roldan

A fundamental principle of learning theory is that there is a trade-off between the complexity of a prediction rule and its ability to generalize. Modern machine learning models do not obey this paradigm: They produce an accurate prediction…

Machine Learning · Computer Science 2021-06-18 Koby Bibas , Meir Feder

Monte Carlo maximum likelihood (MCML) provides an elegant approach to find maximum likelihood estimators (MLEs) for latent variable models. However, MCML algorithms are computationally expensive when the latent variables are…

Computation · Statistics 2020-08-05 Jaewoo Park , Murali Haran

Cluster-weighted modeling (CWM) is a mixture approach for modeling the joint probability of a response variable and a set of explanatory variables. The parameters are estimated by means of the expectation-maximization algorithm according to…

Computation · Statistics 2013-08-09 Salvatore Ingrassia , Simona C. Minotti

Empirical economic research frequently applies maximum likelihood estimation in cases where the likelihood function is analytically intractable. Most of the theoretical literature focuses on maximum simulated likelihood (MSL) estimators,…

Econometrics · Economics 2019-08-13 Michael Griebel , Florian Heiss , Jens Oettershagen , Constantin Weiser

Mixture modelling involves explaining some observed evidence using a combination of probability distributions. The crux of the problem is the inference of an optimal number of mixture components and their corresponding parameters. This…

Machine Learning · Computer Science 2015-03-02 Parthan Kasarapu , Lloyd Allison

Score matching is an alternative to maximum likelihood (ML) for estimating a probability distribution parametrized up to a constant of proportionality. By fitting the ''score'' of the distribution, it sidesteps the need to compute this…

Machine Learning · Computer Science 2023-06-06 Chirag Pabbaraju , Dhruv Rohatgi , Anish Sevekari , Holden Lee , Ankur Moitra , Andrej Risteski

This paper proposes and axiomatizes a new updating rule: Relative Maximum Likelihood (RML) for ambiguous beliefs represented by a set of priors (C). This rule takes the form of applying Bayes' rule to a subset of C. This subset is a linear…

Theoretical Economics · Economics 2024-10-15 Xiaoyu Cheng

Maximum regularized likelihood estimators (MRLEs) are arguably the most established class of estimators in high-dimensional statistics. In this paper, we derive guarantees for MRLEs in Kullback-Leibler divergence, a general measure of…

Machine Learning · Statistics 2018-10-18 Rui Zhuang , Johannes Lederer

Maximum likelihood (ML) estimation using Newton's method in nonlinear state space models (SSMs) is a challenging problem due to the analytical intractability of the log-likelihood and its gradient and Hessian. We estimate the gradient and…

Computation · Statistics 2016-03-11 Manon Kok , Johan Dahlin , Thomas B. Schön , Adrian Wills

The marginal likelihood is a well established model selection criterion in Bayesian statistics. It also allows to efficiently calculate the marginal posterior model probabilities that can be used for Bayesian model averaging of quantities…

Computation · Statistics 2016-11-07 Aliaksandr Hubin , Geir Storvik

Posterior sampling for high-dimensional Bayesian inverse problems is a common challenge in real-world applications. Randomized Maximum Likelihood (RML) is an optimization based methodology that gives samples from an approximation to the…

Computation · Statistics 2024-09-05 Valentin Breaz , Richard Wilkinson

We introduce a parameterization method called Neural Bayes which allows computing statistical quantities that are in general difficult to compute and opens avenues for formulating new objectives for unsupervised representation learning.…

Machine Learning · Statistics 2020-02-24 Devansh Arpit , Huan Wang , Caiming Xiong , Richard Socher , Yoshua Bengio