English
Related papers

Related papers: On redundancy of memoryless sources over countable…

200 papers

What are the distinct ways in which a set of predictor variables can provide information about a target variable? When does a variable provide unique information, when do variables share redundant information, and when do variables combine…

Information Theory · Computer Science 2018-05-04 Conor Finn , Joseph T Lizier

A general method of source coding over expansion is proposed in this paper, which enables one to reduce the problem of compressing an analog (continuous-valued source) to a set of much simpler problems, compressing discrete sources.…

Information Theory · Computer Science 2013-08-13 Hongbo Si , O. Ozan Koyluoglu , Sriram Vishwanath

We study and propose schemes that map messages onto constant-weight codewords using variable-length prefixes. We provide polynomial-time computable formulas that estimate the average number of redundant bits incurred by our schemes. In…

Information Theory · Computer Science 2023-07-04 Duc Tu Dao , Han Mao Kiah , Tuan Thanh Nguyen

The problem of guessing a random string is revisited. A close relation between guessing and compression is first established. Then it is shown that if the sequence of distributions of the information spectrum satisfies the large deviation…

Information Theory · Computer Science 2010-08-12 Manjesh Kumar Hanawal , Rajesh Sundaresan

The interactions between three or more random variables are often nontrivial, poorly understood, and yet, are paramount for future advances in fields such as network information theory, neuroscience, genetics and many others. In this work,…

Information Theory · Computer Science 2016-04-20 Fernando Rosas , Vasilis Ntranos , Christopher J. Ellison , Sofie Pollin , Marian Verhelst

In a system of three stochastic variables, the Partial Information Decomposition (PID) of Williams and Beer dissects the information that two variables (sources) carry about a third variable (target) into nonnegative information atoms that…

Information Theory · Computer Science 2017-08-30 Giuseppe Pica , Eugenio Piasini , Daniel Chicharro , Stefano Panzeri

Given a sufficient statistic for a parametric family of distributions, one can estimate the parameter without access to the data. However, the memory or code size for storing the sufficient statistic may nonetheless still be prohibitive.…

Information Theory · Computer Science 2017-11-17 Masahito Hayashi , Vincent Y. F. Tan

The statistical distribution, when determined from an incomplete set of constraints, is shown to be suitable as host for encrypted information. We design an encoding/decoding scheme to embed such a distribution with hidden information. The…

Statistical Mechanics · Physics 2015-06-25 L. Rebollo-Neira , A Plastino

A random dense countable set is characterized (in distribution) by independence and stationarity. Two examples are `Brownian local minima' and `unordered infinite sample'. They are identically distributed; the former ad hoc proof of this…

Probability · Mathematics 2007-05-23 Boris Tsirelson

We propose two types of universal codes that are suited to two asymptotic regimes when the output alphabet is possibly continuous. The first class has the property that the error probability decays exponentially fast and we identify an…

Information Theory · Computer Science 2024-09-10 Masahito Hayashi

This thesis deals with the problem of communicating and storing non-sequential data. We investigate this problem through the lens of lossless source coding, also sometimes referred to as lossless compression, from both an algorithmic and…

Information Theory · Computer Science 2024-11-25 Daniel Severo

The problem of joint universal source coding and modeling, addressed by Rissanen in the context of lossless codes, is generalized to fixed-rate lossy coding of continuous-alphabet memoryless sources. We show that, for bounded distortion…

Information Theory · Computer Science 2016-11-15 Maxim Raginsky

Universal coding of integers~(UCI) is a class of variable-length code, such that the ratio of the expected codeword length to $\max\{1,H(P)\}$ is within a constant factor, where $H(P)$ is the Shannon entropy of the decreasing probability…

Information Theory · Computer Science 2022-04-18 Wei Yan , Sian-Jheng Lin , Yunghsiang S. Han

We introduce the concept of an extremely negatively dependent (END) sequence of random variables with a given common marginal distribution. The END structure, as a new benchmark for negative dependence, is comparable to comonotonicity and…

Probability · Mathematics 2015-07-28 Bin Wang , Ruodu Wang

A new run length encoding algorithm for lossless data compression that exploits positional redundancy by representing data in a two-dimensional model of concentric circles is presented. This visual transform enables detection of runs (each…

Data Structures and Algorithms · Computer Science 2021-07-30 Pranav Venkatram

For statistical learning, categorical variables in a table are usually considered as discrete entities and encoded separately to feature vectors, e.g., with one-hot encoding. "Dirty" non-curated data gives rise to categorical variables with…

Machine Learning · Computer Science 2018-06-05 Patricio Cerda , Gaël Varoquaux , Balázs Kégl

We present a data structure that stores a sequence $s[1..n]$ over alphabet $[1..\sigma]$ in $n\Ho(s) + o(n)(\Ho(s){+}1)$ bits, where $\Ho(s)$ is the zero-order entropy of $s$. This structure supports the queries \access, \rank\ and \select,…

Data Structures and Algorithms · Computer Science 2012-04-03 Jeremy Barbay , Francisco Claude , Travis Gagie , Gonzalo Navarro , Yakov Nekrich

Probability estimation is essential for every statistical data compression algorithm. In practice probability estimation should be adaptive, recent observations should receive a higher weight than older observations. We present a…

Information Theory · Computer Science 2015-01-12 Christopher Mattern

Storage systems have a strong need for substantially improving their error correction capabilities, especially for long-term storage where the accumulating errors can exceed the decoding threshold of error-correcting codes (ECCs). In this…

Information Theory · Computer Science 2018-11-12 Pulakesh Upadhyaya , Anxiao , Jiang

Encoding data as a set of unordered strings is receiving great attention as it captures one of the basic features of DNA storage systems. However, the challenge of constructing optimal redundancy codes for this channel remained elusive. In…

Information Theory · Computer Science 2023-08-16 Jin Sima , Netanel Raviv , Jehoshua Bruck
‹ Prev 1 8 9 10 Next ›