English
Related papers

Related papers: Co-Occurrence Patterns in the Voynich Manuscript

200 papers

The Voynich manuscript is a medieval book written in an unknown script. This paper studies the relation between similarly spelled words in the Voynich manuscript. By means of a detailed analysis of similar spelled words it was possible to…

Cryptography and Security · Computer Science 2015-12-11 Torsten Timm

While the use of statistical physics methods to analyze large corpora has been useful to unveil many patterns in texts, no comprehensive investigation has been performed investigating the properties of statistical measurements across…

Physics and Society · Physics 2013-07-04 Diego R. Amancio , Eduardo G. Altmann , Diego Rybski , Osvaldo N. Oliveira , Luciano da F. Costa

The statistical properties of letters frequencies in European literature texts are investigated. The determination of logarithmic dependence of letters sequence for one-language and two-language texts are examined. The pare of languages is…

The late medieval Voynich Manuscript (VM) has resisted decryption and was considered a meaningless hoax or an unsolvable cipher. Here, we provide evidence that the VM is written in natural language by establishing a relation of the Voynich…

Computation and Language · Computer Science 2017-10-09 J. Michael Herrmann

This study explores the cryptic Voynich Manuscript, by looking for subtle signs of scribal intent hidden in overlooked features of the "Voynichese" script. The findings indicate that distributions of tokens within paragraphs vary…

Computation and Language · Computer Science 2024-04-23 Andrew Steckley , Noah Steckley

This paper outlines the creation of three corpora for multilingual comparison and analysis of the Voynich manuscript: a corpus of Voynich texts partitioned by Currier language, scribal hand, and transcription system, a corpus of 294…

Computation and Language · Computer Science 2021-05-20 Luke Lindemann , Claire Bowern

A model of co-occurrence in bitext is a boolean predicate that indicates whether a given pair of word tokens co-occur in corresponding regions of the bitext space. Co-occurrence is a precondition for the possibility that two tokens might be…

cmp-lg · Computer Science 2007-05-23 I. Dan Melamed

The Voynich manuscript is the book initially dated as fifteenth century book. It written using specific and smart coding methods. This article describes the methods how it was analyzed and how coding keys were found. The last manuscript…

Cryptography and Security · Computer Science 2018-07-02 Alexander Ulyanenkov

Witnesses of medieval literary texts, preserved in manuscript, are layered objects , being almost exclusively copies of copies. This results in multiple and hard to distinguish linguistic strata -- the author's scripta interacting with the…

Computation and Language · Computer Science 2018-02-06 Jean-Baptiste Camps

We address the problem of predicting similarity between a pair of handwritten document images written by different individuals. This has applications related to matching and mining in image collections containing handwritten content. A…

Computer Vision and Pattern Recognition · Computer Science 2016-05-20 Praveen Krishnan , C. V. Jawahar

This article presents the results of investigations using topic modeling of the Voynich Manuscript (Beinecke MS408). Topic modeling is a set of computational methods which are used to identify clusters of subjects within text. We use latent…

Computation and Language · Computer Science 2021-07-08 Rachel Sterneck , Annie Polish , Claire Bowern

The Voynich Manuscript (VMS) exhibits a script of uncertain origin whose grapheme sequences have resisted linguistic analysis. We present a systematic analysis of its grapheme sequences, revealing two complementary structural layers: a…

Computation and Language · Computer Science 2026-04-23 Christophe Parisel

While the Voynich Manuscript was almost certainly written left-to-right (LTR), the question whether the underlying script or cipher reads LTR or right-to-left (RTL) has received little quantitative attention. We introduce a statistical…

Cryptography and Security · Computer Science 2025-09-25 Christophe Parisel

With the increasing number of texts made available on the Internet, many applications have relied on text mining tools to tackle a diversity of problems. A relevant model to represent texts is the so-called word adjacency (co-occurrence)…

Computation and Language · Computer Science 2019-02-25 Henrique F. de Arruda , Vanessa Q. Marinho , Luciano da F. Costa , Diego R. Amancio

A "monkey book" is a book consisting of a random distribution of letters and blanks, where a group of letters surrounded by two blanks is defined as a word. We compare the statistics of the word distribution for a monkey book with the…

Data Analysis, Statistics and Probability · Physics 2011-09-09 Sebastian Bernhardsson , Seung Ki Baek , Petter Minnhagen

The Voynich MS is an illustrated 15th century manuscript, whose text is written in an unknown alphabet, which has not been translated until today. In 2004 Gordon Rugg published a paper in which he proposed that this text is likely to be…

Computers and Society · Computer Science 2021-04-27 René Zandbergen

In this study, we present a generalizable workflow to identify documents in a historic language with a nonstandard language and script combination, Armeno-Turkish. We introduce the task of detecting distinct patterns of multilinguality…

Computation and Language · Computer Science 2024-01-29 Hale Sirin , Sabrina Li , Tom Lippincott

Due to the availability of references of research papers and the rich information contained in papers, various citation analysis approaches have been proposed to identify similar documents for scholar recommendation. Despite of the success…

Information Retrieval · Computer Science 2017-03-21 Han Tian , Hankz Hankui Zhuo

In this paper, we present the SharedCanvas model for describing the layout of culturally important, hand-written objects such as medieval manuscripts, which is intended to be used as a common input format to presentation interfaces. The…

Digital Libraries · Computer Science 2011-10-18 Robert Sanderson , Hennie Brugman , Benjamin Albritton , Herbert Van de Sompel

The word inference problem is to determine languages such that the information on the number of occurrences of those subwords in the language can uniquely identify a word. A considerable amount of work has been done on this problem, but the…

Combinatorics · Mathematics 2021-10-29 Ghajendran Poovanandran , Jamie Simpson , Wen Chean Teh
‹ Prev 1 2 3 10 Next ›