English
Related papers

Related papers: Longest common subsequences in sets of words

200 papers

Let k be a positive integer. A sequence s over an n-element alphabet A is called a k-radius sequence if every two symbols from A occur in s at distance of at most k. Let f_k(n) denote the length of a shortest k-radius sequence over A. We…

Combinatorics · Mathematics 2011-05-19 Jerzy W. Jaromczyk , Zbigniew Lonc , Miroslaw Truszczynski

Various approaches to alignment-free sequence comparison are based on the length of exact or inexact word matches between two input sequences. Haubold {\em et al.} (2009) showed how the average number of substitutions between two DNA…

Populations and Evolution · Quantitative Biology 2017-09-06 Burkhard Morgenstern , Svenja Schöbel , Chris-André Leimeister

Two same length words are $d$-equivalent if they have same descent set and same underlying alphabet. In particular, two same length permutations are $d$-equivalent if they have same descent set. The popularity of a pattern in a set of words…

Discrete Mathematics · Computer Science 2019-11-13 Jean-Luc Baril , Vincent Vajnovszki

Staircase words are words in which consecutive letters do not differ by more than $1$. We generalize this by extending the restriction to letters lying further apart from each other and obtain the corresponding generating functions, which…

Combinatorics · Mathematics 2025-02-04 Sela Fried

The Longest Common Subsequence (LCS) is the problem of finding a subsequence among a set of strings that has two properties of being common to all and is the longest. The LCS has applications in computational biology and text editing, among…

Artificial Intelligence · Computer Science 2023-06-07 Alireza Abdi , Masih Hajsaeedi , Mohsen Hooshmand

A problem of reconstructing words from their subwords involves determining the minimum amount of information needed, such as multisets of scattered subwords of a specific length or the frequency of scattered subwords from a given set, in…

Discrete Mathematics · Computer Science 2025-12-04 Sergey Luchinin , Svetlana Puzynina , Michaël Rao

Longest Common Subsequence ($LCS$) deals with the problem of measuring similarity of two strings. While this problem has been analyzed for decades, the recent interest stems from a practical observation that considering single characters is…

Data Structures and Algorithms · Computer Science 2018-05-25 Filip Pavetić , Ivan Katanić , Gustav Matula , Goran Žužić , Mile Šikić

We extract brilliant ideas of Sandi Klavzar, Michel Mollard, and Marko Petkovsek who used them to solve one very specific enumeration problem, namely counting the number of words in the alphabet {0,1} of length n avoiding two consecutive…

Combinatorics · Mathematics 2023-04-25 Shalosh B. Ekhad , Doron Zeilberger

We give an exposition of Schensted's algorithm to find the length of the longest increasing subword of a word in an ordered alphabet, and Greene's generalization of Schensted's results using Knuth equivalence. We announce a generalization…

Combinatorics · Mathematics 2018-11-07 Amritanshu Prasad

We obtain an explicit formula for the variance of the number of $k$-peaks in a uniformly random permutation. This is then used to obtain an asymptotic formula for the variance of the length of longest $k$-alternating subsequence in random…

Probability · Mathematics 2026-04-15 Recep Altar Çiçeksiz , Yunus Emre Demirci , Ümit Işlak

A generalization of the well--known Fibonacci sequence is the $k$--Fibonacci sequence with some fixed integer $k\ge 2$. The first $k$ terms of this sequence are $0,\ldots,0,1$, and each term afterwards is the sum of the preceding $k$ terms.…

Number Theory · Mathematics 2020-08-25 Eric F. Bravo , Jhon J. Bravo , Carlos A. Gómez

A {\it superpattern} is a string of characters of length $n$ that contains as a subsequence, and in a sense that depends on the context, all the smaller strings of length $k$ in a certain class. We prove structural and probabilistic results…

Combinatorics · Mathematics 2016-03-08 Yonah Biers-Ariel , Yiguang Zhang , Anant Godbole

The locality of words is a relatively young structural complexity measure, introduced by Day et al. in 2017 in order to define classes of patterns with variables which can be matched in polynomial time. The main tool used to compute the…

Formal Languages and Automata Theory · Computer Science 2020-08-18 Pamela Fleischmann , Lukas Haschke , Florin Manea , Dirk Nowotka , Cedric Tsatia Tsida , Judith Wiedenbeck

The longest common substring with $k$-mismatches problem is to find, given two strings $S_1$ and $S_2$, a longest substring $A_1$ of $S_1$ and $A_2$ of $S_2$ such that the Hamming distance between $A_1$ and $A_2$ is $\le k$. We introduce a…

Data Structures and Algorithms · Computer Science 2015-04-08 Tomas Flouri , Emanuele Giaquinta , Kassian Kobert , Esko Ukkonen

Fici, Restivo, Silva, and Zamboni define a $\textit{$k$-anti-power}$ to be a concatenation of $k$ consecutive words that are pairwise distinct and have the same length. They ask for the maximum $k$ such that every aperiodic recurrent word…

Combinatorics · Mathematics 2019-02-05 Aaron Berger , Colin Defant

Two words are $k$-binomially equivalent if each subword of length at most $k$ occurs the same number of times in both words. The $k$-binomial complexity of an infinite word is a counting function that maps $n$ to the number of $k$-binomial…

Combinatorics · Mathematics 2022-12-07 Michel Rigo , Manon Stipulanti , Markus A. Whiteland

We show that the number of length-n words over a k-letter alphabet having no even palindromic prefix is the same as the number of length-n unbordered words, by constructing an explicit bijection between the two sets. A slightly different…

Discrete Mathematics · Computer Science 2020-06-05 Daniel Gabric , Jeffrey Shallit

A $d$-subsequence of a sequence $\varphi = x_1\dots x_n$ is a subsequence $x_i x_{i+d} x_{i+2d} \dots$, for any positive integer $d$ and any $i$, $1 \le i \le n$. A \textit{$k$-Thue sequence} is a sequence in which every $d$-subsequence,…

Combinatorics · Mathematics 2020-05-15 Borut Lužar , Martina Mockovčiaková , Pascal Ochem , Alexandre Pinlou , Roman Soták

A \emph{power} is a word of the form $\underbrace{uu...u}_{k \; \text{times}}$, where $u$ is a word and $k$ is a positive integer; the power is also called a {\em $k$-power} and $k$ is its {\em exponent}. We prove that for any $k \ge 2$,…

Combinatorics · Mathematics 2022-05-23 Shuo Li , Jakub Pachocki , Jakub Radoszewski

Consider a string of $n$ positions, i.e. a discrete string of length $n$. Units of length $k$ are placed at random on this string in such a way that they do not overlap, and as often as possible, i.e. until all spacings between neighboring…

Probability · Mathematics 2007-05-23 Chris A. J. Klaassen , J. Theo Runnenburg