English
Related papers

Related papers: Redundancy of information: lowering dimension

200 papers

We show how to calculate the finite-state dimension (equivalently, the finite-state compressibility) of a saturated sets $X$ consisting of {\em all} infinite sequences $S$ over a finite alphabet $\Sigma_m$ satisfying some given condition…

Computational Complexity · Computer Science 2007-05-23 Xiaoyang Gu , Jack H. Lutz

The size $b$ of the smallest bidirectional macro scheme, which is arguably the most general copy-paste scheme to generate a given sequence, is considered to be the strictest reachable measure of repetitiveness. It is strictly lower-bounded…

Data Structures and Algorithms · Computer Science 2021-05-31 Gonzalo Navarro , Cristian Urbina

A descent $k$ of a permutation $\pi=\pi_{1}\pi_{2}\dots\pi_{n}$ is called a big descent if $\pi_{k}>\pi_{k+1}+1$; denote the number of big descents of $\pi$ by $\operatorname{bdes}(\pi)$. We study the distribution of the…

Combinatorics · Mathematics 2024-09-02 Sergi Elizalde , Johnny Rivera , Yan Zhuang

The d-neighborhood of a word W in the Levenshtein distance is the set of all words at distance at most d from W. Generating the neighborhood of a word W, or related sets of words such as the condensed neighborhood or the super-condensed…

Combinatorics · Mathematics 2025-09-05 Cedric Chauve , Louxin Zhang

High-dimensional classification is a fundamentally important research problem in high-dimensional data analysis. In this paper, we derive a nonasymptotic rate for the minimax excess misclassification risk when feature dimension…

Statistics Theory · Mathematics 2023-03-07 Shuoyang Wang , Zuofeng Shang

The minimum average number of bits need to describe a random variable is its entropy, assuming knowledge of the underlying statistics On the other hand, universal compression supposes that the distribution of the random variable, while…

Information Theory · Computer Science 2014-04-02 Maryam Hosseini , Narayana Santhanam

Let $\psi:\mathbb{N} \to [0,\infty)$, $\psi(q)=q^{-(1+\tau)}$ and let $\psi$-badly approximable points be those vectors in $\mathbb{R}^{d}$ that are $\psi$-well approximable, but not $c\psi$-well approximable for arbitrarily small constants…

Number Theory · Mathematics 2023-10-04 Henna Koivusalo , Jason Levesley , Benjamin Ward , Xintian Zhang

In this work we review and derive some elementary properties of the discrete renewal sequences based on a positive, finite and integer-valued random variable. Our results consider these sequences as dependent on the probability masses of…

Probability · Mathematics 2024-05-28 Nikolai Nikolov , Mladen Savov

For a finite alphabet $\mathcal{A}$ and a sequence $x \in \mathcal{A}^{\mathbb{N}}$, Kamae and Zamboni defined the maximal pattern complexity function $p^*_x(n)$ as a natural generalization of usual word complexity. They defined a…

Dynamical Systems · Mathematics 2025-08-20 Anh N. Le , Ronnie Pavlov , Casey Schlortt

The nature of the alignment with gaps corresponding to a longest common subsequence (LCS) of two independent iid random sequences drawn from a finite alphabet is investigated. It is shown that such an optimal alignment typically matches…

Probability · Mathematics 2016-04-22 C. Houdré , H. Matzinger

Consider the case where consecutive blocks of N letters of a semi-infinite individual sequence X over a finite-alphabet are being compressed into binary sequences by some one-to-one mapping. No a-priori information about X is available at…

Information Theory · Computer Science 2013-01-25 Jacob Ziv

We study the extreme $L_p$ discrepancy of infinite sequences in the $d$-dimensional unit cube, which uses arbitrary sub-intervals of the unit cube as test sets. This is in contrast to the classical star $L_p$ discrepancy, which uses…

Number Theory · Mathematics 2021-09-15 Ralph Kritzinger , Friedrich Pillichshammer

The information in an individual finite object (like a binary string) is commonly measured by its Kolmogorov complexity. One can divide that information into two parts: the information accounting for the useful regularity present in the…

Computational Complexity · Computer Science 2007-05-23 Paul Vitanyi

We study the entropy $S$ of longest increasing subsequences (LIS), i.e., the logarithm of the number of distinct LIS. We consider two ensembles of sequences, namely random permutations of integers and sequences drawn i.i.d.\ from a limited…

Disordered Systems and Neural Networks · Physics 2020-06-09 Phil Krabbe , Hendrik Schawe , Alexander K. Hartmann

Ochem, Rampersad, and Shallit gave various examples of infinite words avoiding what they called approximate repetitions. An approximate repetition is a factor of the form xx', where x and x' are close to being identical. In their work, they…

Combinatorics · Mathematics 2016-08-03 Serina Camungol , Narad Rampersad

We consider approximation of diameter of a set $S$ of $n$ points in dimension $m$. E$\tilde{g}$ecio$\tilde{g}$lu and Kalantari \cite{kal} have shown that given any $p \in S$, by computing its farthest in $S$, say $q$, and in turn the…

Computational Geometry · Computer Science 2014-10-09 Sharareh Alipour , Bahman Kalantari , Hamid Homapour

Given a sequence composed of a limit number of characters, we try to "read" it as a "text". This involves to segment the sequence into "words". The difficulty is to distinguish good segmentation from enormous number of random ones.Aiming at…

Biological Physics · Physics 2009-11-06 Bin Wang

Let $\mathcal{A}$ be a sequence of $rk$ terms which is made up of $k$ distinct integers each appearing exactly $r$ times in $\mathcal{A}$. The sum of all terms of a subsequence of $\mathcal{A}$ is called a subsequence sum of $\mathcal{A}$.…

Number Theory · Mathematics 2022-11-24 Jagannath Bhanja , Ram Krishna Pandey

We prove an exponential size separation between depth 2 and depth 3 neural networks (with real inputs), when approximating a $\mathcal{O}(1)$-Lipschitz target function to constant accuracy, with respect to a distribution with support in the…

Machine Learning · Computer Science 2024-11-08 Itay Safran , Daniel Reichman , Paul Valiant

It is shown that for two large subclasses of discrete-time nonlinear systems - analytic systems defined on a compact state space and rational systems - the minimum length $r^*$ for input sequences, called here accessibility index of the…

Systems and Control · Computer Science 2019-06-26 Mohammad Amin Sarafrazi , Ewa Pawluszewicz , Zbigniew Bartosiewicz , Ülle Kotta