中文
相关论文

相关论文: Words that almost commute

200 篇论文

Fraenkel and Simpson showed that the number of distinct squares in a word of length n is bounded from above by 2n, since at most two distinct squares have their rightmost, or last, occurrence begin at each position. Improvements by Ilie to…

形式语言与自动机理论 · 计算机科学 2017-08-23 F. Blanchet-Sadri , S. Osborne

We regard a finite word $u=u_1u_2\cdots u_n$ up to word isomorphism as an equivalence relation on $\{1,2,\ldots, n\}$ where $i$ is equivalent to $j$ if and only if $x_i=x_j.$ Some finite words (in particular all binary words) are generated…

组合数学 · 数学 2014-04-04 Tero Harju , Mari Huova , L. Q. Zamboni

How long can a word be that avoids the unavoidable? Word $W$ encounters word $V$ provided there is a homomorphism $\phi$ defined by mapping letters to nonempty words such that $\phi(V)$ is a subword of $W$. Otherwise, $W$ is said to avoid…

组合数学 · 数学 2014-10-30 Joshua Cooper , Danny Rorabaugh

Twins in a finite word are formed by a pair of identical subwords placed at disjoint sets of positions. We investigate the maximum length of twins in a random word over a $k$-letter alphabet. The obtained lower bounds for small values of…

组合数学 · 数学 2023-04-25 Andrzej Dudek , Jarosław Grytczuk , Andrzej Ruciński

In this paper we compare two finite words $u$ and $v$ by the lexicographical order of the infinite words $u^\omega$ and $v^\omega$. Informally, we say that we compare $u$ and $v$ by the infinite order. We show several properties of Lyndon…

离散数学 · 计算机科学 2019-04-02 Francesco Dolce , Antonio Restivo , Christophe Reutenauer

The problem addressed concerns the determination of the average number of successive attempts of guessing a word of a certain length consisting of letters with given probabilities of occurrence. Both first- and second-order approximations…

信息论 · 计算机科学 2015-06-19 Kerstin Andersson

The problem of comparing two bodies of text and searching for words that differ in their usage between them arises often in digital humanities and computational social science. This is commonly approached by training word embeddings on each…

计算与语言 · 计算机科学 2021-12-30 Hila Gonen , Ganesh Jawahar , Djamé Seddah , Yoav Goldberg

The Ulam distance of two permutations on $[n]$ is $n$ minus the length of their longest common subsequence. In this paper, we show that for every $\varepsilon>0$, there exists some $\alpha>0$, and an infinite set $\Gamma\subseteq…

信息论 · 计算机科学 2024-05-14 Elazar Goldenberg , Mursalin Habib , Karthik C. S

Suppose Alice and Bob receive strings $X=(X_1,...,X_n)$ and $Y=(Y_1,...,Y_n)$ each uniformly random in $[s]^n$ but so that $X$ and $Y$ are correlated . For each symbol $i$, we have that $Y_i = X_i$ with probability $1-\eps$ and otherwise…

信息论 · 计算机科学 2012-08-30 Siu On Chan , Elchanan Mossel , Joe Neeman

Recent works on word representations mostly rely on predictive models. Distributed word representations (aka word embeddings) are trained to optimally predict the contexts in which the corresponding words tend to appear. Such models have…

计算与语言 · 计算机科学 2015-04-10 Rémi Lebret , Ronan Collobert

Commability is the finest equivalence relation between locally compact groups such that $G$ and $H$ are equivalent whenever there is a continuous proper homomorphism $G \to H$ with cocompact image. Answering a question of Cornulier, we show…

群论 · 数学 2014-12-18 Mathieu Carette

Given two functions $\mathbf{a}\!:\! [n] \rightarrow [n]$ and $\mathbf{b}\!:\! [n] \rightarrow [n]$ chosen uniformly at random, any word $w=w_1w_2\dots w_k\in \{a,b\}^k$ induces a random function $\mathbf{w}\!:\! [n] \rightarrow [n]$ by…

概率论 · 数学 2026-04-01 Guillaume Chapuy , Guillem Perarnau

We present a report from a series of experiments involving computation of the shortest reset words for automata with small number of states. We confirm that the \v{C}ern\'{y} conjecture is true for all automata with at most 11 states on 2…

形式语言与自动机理论 · 计算机科学 2013-01-11 Jakub Kowalski , Marek Szykuła

The separating words problem asks for the size of the smallest DFA needed to distinguish between two words of length <= n (by accepting one and rejecting the other). In this paper we survey what is known and unknown about the problem,…

形式语言与自动机理论 · 计算机科学 2011-03-24 Erik D. Demaine , Sarah Eisenstat , Jeffrey Shallit , David A. Wilson

A binary word is a map W : N --> {0,1}, and the set of factors of W with length n is F_n(W):={(W(i),W(i+1),...,W(i+n-1)) : i >= 0}. A word is Sturmian if |F_n(W)|=n+1 for every n>0. We show that the sum of the heights (also known as hamming…

组合数学 · 数学 2007-05-23 Kevin O'Bryant

Let M(n, d) be the maximum size of a permutation array on n symbols with pairwise Hamming distance at least d. We use various combinatorial, algebraic, and computational methods to improve lower bounds for M(n, d). We compute the Hamming…

离散数学 · 计算机科学 2016-10-03 Sergey Bereg , Avi Levy , I. Hal Sudborough

A key principle in assessing textual similarity is measuring the degree of semantic overlap between two texts by considering the word alignment. Such alignment-based approaches are intuitive and interpretable; however, they are empirically…

计算与语言 · 计算机科学 2020-11-17 Sho Yokoi , Ryo Takahashi , Reina Akama , Jun Suzuki , Kentaro Inui

In this work, we consider the problem of synchronizing two sets of data where the size of the symmetric difference between the sets is small and, in addition, the elements in the symmetric difference are related through the Hamming distance…

信息论 · 计算机科学 2018-09-14 Ryan Gabrys , Farzad Farnoud

We consider the expected length of the longest common subsequence between two random words of lengths $n$ and $(1-\varepsilon)kn$ over $k$-symbol alphabet. It is well-known that this quantity is asymptotic to $\gamma_{k,\varepsilon} n$ for…

概率论 · 数学 2021-08-20 Boris Bukh , Zichao Dong

Alice and Bob want to know if two strings of length n are almost equal. That is, do they differ on \textit{at most} a bits? Let 0\leq a\leq n-1. We show that any deterministic protocol, as well as any error-free quantum protocol (C*…

计算复杂性 · 计算机科学 2015-01-13 Andris Ambainis , William Gasarch , Aravind Srinavasan , Andrey Utis