English
Related papers

Related papers: Algorithms for Estimating Information Distance wit…

200 papers

We address the problem, not of the determination -- which usually needs numerical methods -- but of an accurate analytical estimation of the distance of a raw elasticity tensor to cubic symmetry and to orthotropy. We point out that there…

Classical Physics · Physics 2022-06-08 Rodrigue Desmorat , Boris Kolev

In this paper, we propose a new protocol for a data compression task, blind quantum data compression, with finite local approximations. The rate of blind data compression is susceptible to approximations even when the approximations are…

Quantum Physics · Physics 2023-06-07 Kohdai Kuroiwa , Debbie Leung

We consider the problem of inferring the probability distribution associated with a language, given data consisting of an infinite sequence of elements of the languge. We do this under two assumptions on the algorithms concerned: (i) like a…

Machine Learning · Computer Science 2014-07-16 Paul M. B. Vitanyi , Nick Chater

Many quantum chemical similarity measures have been derived and substantiated by applying concepts and quantities from information theory to the electron density. To justify the use of information theory, the electron density is usually…

Chemical Physics · Physics 2024-10-14 Guillaume Acke , Daria Van Hende , Ruben Van der Stichelen , Patrick Bultinck

Information bottleneck (IB) is a paradigm to extract information in one target random variable from another relevant random variable, which has aroused great interest due to its potential to explain deep neural networks in terms of…

Information Theory · Computer Science 2023-08-23 Lingyi Chen , Shitong Wu , Wenhao Ye , Huihui Wu , Hao Wu , Wenyi Zhang , Bo Bai , Yining Sun

Optimizing distributed learning systems is an art of balancing between computation and communication. There have been two lines of research that try to deal with slower networks: {\em communication compression} for low bandwidth networks,…

Machine Learning · Computer Science 2019-02-04 Hanlin Tang , Shaoduo Gan , Ce Zhang , Tong Zhang , Ji Liu

F.Giroire has recently proposed an algorithm which returns the approximate number of distincts elements in a large sequence of words, under strong constraints coming from the analysis of large data bases. His estimation is based on…

Statistics Theory · Mathematics 2011-04-25 Philippe Chassaing , Lucas Gerin

The main approach of traditional information retrieval (IR) is to examine how many words from a query appear in a document. A drawback of this approach, however, is that it may fail to detect relevant documents where no or only few words…

Computation and Language · Computer Science 2017-10-19 Sun Kim , Nicolas Fiorini , W. John Wilbur , Zhiyong Lu

Estimating the Shannon entropy of a discrete distribution from which we have only observed a small sample is challenging. Estimating other information-theoretic metrics, such as the Kullback-Leibler divergence between two sparsely sampled…

Data Analysis, Statistics and Probability · Physics 2023-02-24 Angelo Piga , Lluc Font-Pomarol , Marta Sales-Pardo , Roger Guimerà

We design a quantum method for classical information compression that exploits the hidden subgroup quantum algorithm. We consider sequence data in a database with a priori unknown symmetries of the hidden subgroup type. We prove that data…

Quantum Physics · Physics 2024-08-14 Feiyang Liu , Kaiming Bian , Fei Meng , Wen Zhang , Oscar Dahlsten

This paper examines information-theoretic questions regarding the difficulty of compressing data versus the difficulty of decompressing data and the role that information loss plays in this interaction. Finite-state compression and…

Computational Complexity · Computer Science 2009-09-29 David Doty , Philippe Moser

The paper introduces a new lossless, highly robust compression algorithm that similar with LZW algorithm, yet the algorithm discards dictionary processing and uses irregular sequences with massive, random information instead. Then the paper…

Signal Processing · Electrical Eng. & Systems 2020-06-24 Rui Zhu

From a machine learning point of view, identifying a subset of relevant features from a real data set can be useful to improve the results achieved by classification methods and to reduce their time and space complexity. To achieve this…

Machine Learning · Computer Science 2017-05-23 Pietro Cassara , Alessandro Rozza , Mirco Nanni

We propose a general algorithm for approximating nonstandard Bayesian posterior distributions. The algorithm minimizes the Kullback-Leibler divergence of an approximating distribution to the intractable posterior distribution. Our method…

Computation · Statistics 2014-07-29 Tim Salimans , David A. Knowles

Based on information theory, we present a method to determine an optimal Markov approximation for modelling and prediction from time series data. The method finds a balance between minimal modelling errors by taking as much as possible…

Chaotic Dynamics · Physics 2013-05-29 Detlef Holstein , Holger Kantz

The maximum entropy principle is a powerful tool for solving underdetermined inverse problems. This paper considers the problem of discretizing a continuous distribution, which arises in various applied fields. We obtain the approximating…

Numerical Analysis · Mathematics 2020-08-05 Ken'ichiro Tanaka , Alexis Akira Toda

Message identification (M-I) divergence is an important measure of the information distance between probability distributions, similar to Kullback-Leibler (K-L) and Renyi divergence. In fact, M-I divergence with a variable parameter can…

Information Theory · Computer Science 2024-04-08 Rui She , Shanyun Liu , Pingyi Fan

We survey diverse approaches to the notion of information: from Shannon entropy to Kolmogorov complexity. Two of the main applications of Kolmogorov complexity are presented: randomness and classification. The survey is divided in two parts…

Logic in Computer Science · Computer Science 2010-10-20 Marie Ferbus-Zanda , Serge Grigorieff

Meta-analytic methods tend to take all-or-nothing approaches to study-level heterogeneity, assuming all studies are heterogeneous or homogeneous, leading to inefficiency and/or bias in estimation and inference. In this paper, we develop a…

Methodology · Statistics 2026-03-12 Elizabeth M. Davis , Emily C. Hector

This work proposes an algorithm to bound the minimum distance between points on trajectories of a dynamical system and points on an unsafe set. Prior work on certifying safety of trajectories includes barrier and density methods, which do…

Optimization and Control · Mathematics 2023-06-16 Jared Miller , Mario Sznaier
‹ Prev 1 8 9 10 Next ›