中文
相关论文

相关论文: Prefix Codes: Equiprobable Words, Unequal Letter C…

200 篇论文

Traditional error-correcting codes (ECCs) assume a fixed message length, but many scenarios involve ongoing or indefinite transmissions where the message length is not known in advance. For example, when streaming a video, the user should…

数据结构与算法 · 计算机科学 2025-04-09 Klim Efremenko , Or Zamir

A text written using symbols from a given alphabet can be compressed using the Huffman code, which minimizes the length of the encoded text. It is necessary, however, to employ a text-specific codebook, i.e. the symbol-codeword dictionary,…

信息论 · 计算机科学 2022-08-02 Armen E. Allahverdyan , Andranik Khachatryan

Variable-length splittable codes are derived from encoding sequences of ordered integer pairs, where one of the pair's components is upper bounded by some constant, and the other one is any positive integer. Each pair is encoded by the…

信息论 · 计算机科学 2015-08-07 Anatoly V. Anisimov , Igor O. Zavadskyi

We consider coding schemes for computationally bounded channels, which can introduce an arbitrary set of errors as long as (a) the fraction of errors is bounded with high probability by a parameter $p$ and (b) the process which adds the…

信息论 · 计算机科学 2013-03-01 Venkatesan Guruswami , Adam Smith

We relate the computational complexity of finite strings to universal representations of their underlying symmetries. First, Boolean functions are classified using the universal covering topologies of the circuits which enumerate them. A…

信息论 · 计算机科学 2011-09-20 John Scoville

The optimal prefix-free machine U is a universal decoding algorithm used to define the notion of program-size complexity H(s) for a finite binary string s. Since the set of all halting inputs for U is chosen to form a prefix-free set, the…

信息论 · 计算机科学 2016-11-15 Kohtaro Tadaki

We present a linear time and space algorithm computing the leftmost critical factorization of a given string on an unordered alphabet.

数据结构与算法 · 计算机科学 2016-04-12 Dmitry Kosolobov

We design a heuristic method, a genetic algorithm, for the computation of an upper bound of the minimum distance of a linear code over a finite field. By the use of the row reduced echelon form, we obtain a permutation encoding of the…

信息论 · 计算机科学 2018-07-20 José Gómez-Torrecillas , F. J. Lobillo , Gabriel Navarro

AIFV (almost instantaneous fixed-to-variable length) codes are noiseless source codes that can attain a shorter average codeword length than Huffman codes by allowing a time-variant encoder with two code tables and a decoding delay of at…

信息论 · 计算机科学 2023-06-19 Kengo Hashimoto , Ken-ichi Iwata

We consider universal variable-to-fixed length compression of memoryless sources with a fidelity criterion. We design a dictionary codebook over the reproduction alphabet which is used to parse the source stream. Once a source subsequence…

信息论 · 计算机科学 2022-11-24 Nematollah Iri

In this paper we consider block languages, namely sets of words having the same length, and study the deterministic and nondeterministic state complexity of several operations on these languages. Being a subclass of finite languages, the…

形式语言与自动机理论 · 计算机科学 2024-09-12 Guilherme Duarte , Nelma Moreira , Luca Prigioniero , Rogério Reis

Subword tokenization is the de facto standard for tokenization in neural language models and machine translation systems. Three advantages are frequently cited in favor of subwords: shorter encoding of frequent tokens, compositionality of…

计算与语言 · 计算机科学 2024-01-15 Benoist Wolleb , Romain Silvestri , Giorgos Vernikos , Ljiljana Dolamic , Andrei Popescu-Belis

Linear computation coding is concerned with the compression of multidimensional linear functions, i.e. with reducing the computational effort of multiplying an arbitrary vector to an arbitrary, but known, constant matrix. This paper…

信息论 · 计算机科学 2025-07-02 Hans Rosenberger , Johanna S. Fröhlich , Ali Bereyhi , Ralf R. Müller

New bounds on the cardinality of permutation codes equipped with the Ulam distance are presented. First, an integer-programming upper bound is derived, which improves on the Singleton-type upper bound in the literature for some lengths.…

信息论 · 计算机科学 2015-04-21 Faruk Göloğlu , Jüri Lember , Ago-Erik Riet , Vitaly Skachek

We derive a single-letter upper bound to the mismatched-decoding capacity for discrete memoryless channels. The bound is expressed as the mutual information of a transformation of the channel, such that a maximum-likelihood decoding error…

信息论 · 计算机科学 2021-02-16 Ehsan Asadi Kangarshahi , Albert Guillén i Fàbregas

The set of finite words over a well-quasi-ordered set is itself well-quasi-ordered. This seminal result by Higman is a cornerstone of the theory of well-quasi-orderings and has found numerous applications in computer science. However, this…

形式语言与自动机理论 · 计算机科学 2025-01-14 Nathan Lhote , Aliaume Lopez , Lia Schütze

Locally repairable codes (LRC) have recently been a subject of intense research due to theoretical appeal and their application in distributed storage systems. In an LRC, any coordinate of a codeword can be recovered by accessing only few…

信息论 · 计算机科学 2016-07-29 Abhishek Agarwal , Arya Mazumdar

A common complaint about adaptive prefix coding is that it is much slower than static prefix coding. Karpinski and Nekrich recently took an important step towards resolving this: they gave an adaptive Shannon coding algorithm that encodes…

信息论 · 计算机科学 2008-12-18 Travis Gagie , Yakov Nekrich

We introduce a binary embedding framework, called Proximity Preserving Code (PPC), which learns similarity and dissimilarity between data points to create a compact and affinity-preserving binary code. This code can be used to apply fast…

机器学习 · 计算机科学 2020-02-06 Inbal Lav , Shai Avidan , Yoram Singer , Yacov Hel-Or

In this paper we study the redundancy of Huffman codes. In particular, we consider sources for which the probability of one of the source symbols is known. We prove a conjecture of Ye and Yeung regarding the upper bound on the redundancy of…

信息论 · 计算机科学 2016-11-17 Soheil Mohajer , Payam Pakzad , Ali Kakhbod