English
Related papers

Related papers: Maximum Entropy Based Significance of Itemsets

200 papers

We consider the task of estimating a conditional density using i.i.d. samples from a joint distribution, which is a fundamental problem with applications in both classification and uncertainty quantification for regression. For joint…

Statistics Theory · Mathematics 2023-06-16 Blair Bilodeau , Dylan J. Foster , Daniel M. Roy

This paper applies the recently axiomatized Optimum Information Principle (minimize the Kullback-Leibler information subject to all relevant information) to nonparametric density estimation, which provides a theoretical foundation as well…

Statistics Theory · Mathematics 2011-03-28 Alexis Akira Toda

We present evidence that the best model for empirical volume-price distributions is not always the same and it strongly depends in (i) the region of the volume-price spectrum that one wants to model and (ii) the period in time that is being…

Statistical Finance · Quantitative Finance 2015-06-22 Paulo Rocha , Frank Raischel , João Pedro Boto , Pedro G. Lind

Recently, varextropy has been introduced as a new dispersion index and a measure of information. In this article, we derive the generating function of extropy and present its infinite series representation. Furthermore, we propose new…

Statistics Theory · Mathematics 2025-12-12 Faranak Goodarzi , Somayeh Ghafouri

Distance measures between quantum states like the trace distance and the fidelity can naturally be defined by optimizing a classical distance measure over all measurement statistics that can be obtained from the respective quantum states.…

Quantum Physics · Physics 2017-11-09 Mario Berta , Omar Fawzi , Marco Tomamichel

Automatic detection of anomalies in space- and time-varying measurements is an important tool in several fields, e.g., fraud detection, climate analysis, or healthcare monitoring. We present an algorithm for detecting anomalous regions in…

Machine Learning · Statistics 2019-07-24 Björn Barz , Erik Rodner , Yanira Guanche Garcia , Joachim Denzler

The Information Bottleneck method is a learning technique that seeks a right balance between accuracy and generalization capability through a suitable tradeoff between compression complexity, measured by minimum description length, and…

Information Theory · Computer Science 2020-11-04 Mohammad Mahdi Mahvari , Mari Kobayashi , Abdellatif Zaidi

Variable importance in regression analyses is of considerable interest in a variety of fields. There is no unique method for assessing variable importance. However, a substantial share of the available literature employs Shapley values,…

Methodology · Statistics 2026-01-05 Sinan Acemoglu , Christian Kleiber , Jörg Urban

Calibration methods have been widely studied in survey sampling over the last decades. Viewing calibration as an inverse problem, we extend the calibration technique by using a maximum entropy method. Finding the optimal weights is achieved…

Methodology · Statistics 2009-09-23 Fabrice Gamboa , Jean-Michel Loubes , Paul Rochet

Maximum entropy models provide the least constrained probability distributions that reproduce statistical properties of experimental datasets. In this work we characterize the learning dynamics that maximizes the log-likelihood in the case…

Disordered Systems and Neural Networks · Physics 2016-09-21 Ulisse Ferrari

Given a sequence composed of a limit number of characters, we try to "read" it as a "text". This involves to segment the sequence into "words". The difficulty is to distinguish good segmentation from enormous number of random ones.Aiming at…

Biological Physics · Physics 2009-11-06 Bin Wang

The article presents new sup-sums principles for integral F-divergence for arbitrary convex function F and arbitrary (not necessarily positive and absolutely continuous) measures. As applications of these results we derive the corresponding…

Statistics Theory · Mathematics 2019-09-17 V. I. Bakhtin , A. V. Lebedev

The popularity of transformer-based text embeddings calls for better statistical tools for measuring distributions of such embeddings. One such tool would be a method for ranking texts within a corpus by centrality, i.e. assigning each text…

Computation and Language · Computer Science 2023-10-24 Parker Seegmiller , Sarah Masud Preum

We consider asymptotic distributions of maximum deviations of sample covariance matrices, a fundamental problem in high-dimensional inference of covariances. Under mild dependence conditions on the entries of the data matrices, we establish…

Statistics Theory · Mathematics 2011-09-05 Han Xiao , Wei Biao Wu

How can we assess the reliability of a dataset without access to ground truth? We introduce the problem of reliability scoring for datasets collected from potentially strategic sources. The true data are unobserved, but we see outcomes of…

Machine Learning · Computer Science 2025-10-21 Yiling Chen , Shi Feng , Paul Kattuman , Fang-Yi Yu

In this paper we review various information-theoretic characterizations of the approach to equilibrium in biological systems. The replicator equation, evolutionary game theory, Markov processes and chemical reaction networks all describe…

Information Theory · Computer Science 2017-08-22 John C. Baez , Blake S. Pollard

Importance sampling is widely used in machine learning and statistics, but its power is limited by the restriction of using simple proposals for which the importance weights can be tractably calculated. We address this problem by studying…

Machine Learning · Statistics 2016-10-18 Qiang Liu , Jason D. Lee

Classical change point analysis aims at (1) detecting abrupt changes in the mean of a possibly non-stationary time series and at (2) identifying regions where the mean exhibits a piecewise constant behavior. In many applications however, it…

Statistics Theory · Mathematics 2020-02-17 Axel Bücher , Holger Dette , Florian Heinrichs

We study the excess minimum risk in statistical inference, defined as the difference between the minimum expected loss in estimating a random variable from an observed feature vector and the minimum expected loss in estimating the same…

Information Theory · Computer Science 2023-09-29 László Györfi , Tamás Linder , Harro Walk

Statistical significance measures the reliability of a result obtained from a random experiment. We investigate the number of repetitions needed for a statistical result to have a certain significance. In the first step, we consider…

Methodology · Statistics 2024-06-19 Maike Tormählen , Galiya Klinkova , Michael Grabinski
‹ Prev 1 8 9 10 Next ›