English
Related papers

Related papers: Revisiting enumerative two-part crude MDL for Bern…

200 papers

The Minimum Description Length (MDL) principle is solidly based on a provably ideal method of inference using Kolmogorov complexity. We test how the theory behaves in practice on a general problem in model selection: that of learning the…

Data Analysis, Statistics and Probability · Physics 2007-05-23 Qiong Gao , Ming Li , Paul Vitanyi

This paper introduces a new notion of dimensionality of probabilistic models from an information-theoretic view point. We call it the "descriptive dimension"(Ddim). We show that Ddim coincides with the number of independent parameters for…

Machine Learning · Computer Science 2019-10-28 Kenji Yamanishi

The normalized maximized likelihood (NML) provides the minimax regret solution in universal data compression, gambling, and prediction, and it plays an essential role in the minimum description length (MDL) method of statistical modeling…

Information Theory · Computer Science 2014-01-29 Andrew Barron , Teemu Roos , Kazuho Watanabe

In this correspondence, we focus on the performance analysis of the widely-used minimum description length (MDL) source enumeration technique in array processing. Unfortunately, available theoretical analysis exhibit deviation from the…

Information Theory · Computer Science 2015-05-13 Farzan Haddadi , Mohammadreza Malekmohammadi , Mohammad Mahdi Nayebi , Mohammad Reza Aref

By applying Sklar's theorem to the Multivariate Bernoulli Distribution (MBD), this paper proposes a framework to decouple marginal distributions from the dependence structure, clarifying interactions among binary variables. Explicit…

Methodology · Statistics 2025-09-16 Arturo Erdely

This paper shows that the normalized maximum likelihood~(NML) code-length calculated in [1] is an upper bound on the NML code-length strictly calculated for the Gaussian Mixture Model. When we use this upper bound on the NML code-length, we…

Information Theory · Computer Science 2018-11-20 So Hirai , Kenji Yamanishi

Interpretable classifiers have recently witnessed an increase in attention from the data mining community because they are inherently easier to understand and explain than their more complex counterparts. Examples of interpretable…

Machine Learning · Computer Science 2019-11-01 Hugo M. Proença , Matthijs van Leeuwen

Deep energy-based models (EBMs), which use deep neural networks (DNNs) as energy functions, are receiving increasing attention due to their ability to learn complex distributions. To train deep EBMs, the maximum likelihood estimation (MLE)…

Machine Learning · Computer Science 2022-05-31 Beomsu Kim , Jong Chul Ye

Binary logit (BNL) and multinomial logit (MNL) models are the two most widely used discrete choice models for travel behavior modeling and prediction. However, in many scenarios, the collected data for those models are subject to…

Optimization and Control · Mathematics 2025-06-02 Baichuan Mo , Yunhan Zheng , Xiaotong Guo , Ruoyun Ma , Jinhua Zhao

In this paper, we consider the multivariate Bernoulli distribution as a model to estimate the structure of graphs with binary nodes. This distribution is discussed in the framework of the exponential family, and its statistical properties…

Applications · Statistics 2013-11-13 Bin Dai , Shilin Ding , Grace Wahba

Approaches to bivariate causal discovery based on the minimum description length (MDL) principle approximate the (uncomputable) Kolmogorov complexity of the models in each causal direction, selecting the one with the lower total complexity.…

Machine Learning · Computer Science 2026-04-08 Tiago Brogueira , Mário A. T. Figueiredo

The sum of $n$ {non-independent} Bernoulli random variables could be modeled in several different ways. One of these is the Multiplicative Binomial Distribution (MBD), introduced by Altham (1978) and revised by Lovison (1998). In this work,…

Statistics Theory · Mathematics 2018-02-26 Francesca Fortunato

We present the first theoretical framework that connects predictive coding (PC), a biologically inspired local learning rule, with the minimum description length (MDL) principle in deep networks. We prove that layerwise PC performs…

Machine Learning · Computer Science 2025-07-17 Benjamin Prada , Shion Matsumoto , Abdul Malik Zekri , Ankur Mali

We provide a complete characterization of the entire regularization curve of a modified two-part-code Minimum Description Length (MDL) learning rule for binary classification, based on an arbitrary prior or description language. Grunwald…

Machine Learning · Statistics 2025-03-12 Xiaohan Zhu , Nathan Srebro

Recent advancements in large language models (LLMs) have significantly improved code generation and program comprehension, accelerating the evolution of software engineering. Current methods primarily enhance model performance by leveraging…

Computation and Language · Computer Science 2025-07-04 Weijie Lyu , Sheng-Jun Huang , Xuan Xia

This paper develops nonparametric estimation for discrete choice models based on the mixed multinomial logit (MMNL) model. It has been shown that MMNL models encompass all discrete choice models derived under the assumption of random…

Statistics Theory · Mathematics 2011-02-25 Pierpaolo De Blasi , Lancelot F. James , John W. Lau

Optical and near-IR (NIR) line profiles of many ageing core-collapse supernovae (CCSNe) exhibit an apparently asymmetric bluewards shift often attributed to greater extinction by internal dust of redshifted radiation emitted from the…

Solar and Stellar Astrophysics · Physics 2019-01-09 Antonia Bevan

The modelling of data on a spherical surface requires the consideration of directional probability distributions. To model asymmetrically distributed data on a three-dimensional sphere, Kent distributions are often used. The moment…

Machine Learning · Computer Science 2015-06-29 Parthan Kasarapu

Multi-label classification aims to classify instances with discrete non-exclusive labels. Most approaches on multi-label classification focus on effective adaptation or transformation of existing binary and multi-class learning approaches…

Machine Learning · Computer Science 2019-01-03 Piotr Szymański , Tomasz Kajdanowicz , Nitesh Chawla

Learning the structure of Bayesian networks and causal relationships from observations is a common goal in several areas of science and technology. We show that the prequential minimum description length principle (MDL) can be used to…

Machine Learning · Computer Science 2021-07-13 Jorg Bornschein , Silvia Chiappa , Alan Malek , Rosemary Nan Ke