中文
相关论文

相关论文: Phylogeny and geometry of languages from normalize…

200 篇论文

The idea of measuring distance between languages seems to have its roots in the work of the French explorer Dumont D'Urville \cite{Urv}. He collected comparative words lists of various languages during his voyages aboard the Astrolabe from…

计算与语言 · 计算机科学 2015-05-14 Filippo Petroni , Maurizio Serva

The idea of measuring distance between languages seems to have its roots in the work of the French explorer Dumont D'Urville (D'Urville 1832). He collected comparative words lists of various languages during his voyages aboard the Astrolabe…

计算与语言 · 计算机科学 2009-12-07 Filippo Petroni , Maurizio Serva

Languages evolve over time in a process in which reproduction, mutation and extinction are all possible, similar to what happens to living organisms. Using this similarity it is possible, in principle, to build family trees which show the…

计算与语言 · 计算机科学 2012-07-03 Maurizio Serva

Phylogenetic trees can be reconstructed from the matrix which contains the distances between all pairs of languages in a family. Recently, we proposed a new method which uses normalized Levenshtein distances among words with same meaning…

计算与语言 · 计算机科学 2015-05-14 Filippo Petroni , Maurizio Serva

The evolution of languages closely resembles the evolution of haploid organisms. This similarity has been recently exploited \cite{GA,GJ} to construct language trees. The key point is the definition of a distance among all pairs of…

物理与社会 · 物理学 2009-11-13 Maurizio Serva , Filippo Petroni

This research project aimed to overcome the challenge of analysing human language relationships, facilitate the grouping of languages and formation of genealogical relationship between them by developing automated comparison techniques.…

计算与语言 · 计算机科学 2020-02-03 Gabija Mikulyte , David Gilbert

Lexical resemblances among a group of languages indicate that the languages could be genetically related, i.e., they could have descended from a common ancestral language. However, such resemblances can arise by chance and, hence, need not…

计算与语言 · 计算机科学 2024-04-02 V. S. D. S. Mahesh Akavarapu , Arnab Bhattacharya

Dictionary lookup methods are popular in dealing with ambiguous letters which were not recognized by Optical Character Readers. However, a robust dictionary lookup method can be complex as apriori probability calculation or a large…

信息论 · 计算机科学 2011-01-07 Rishin Haldar , Debajyoti Mukhopadhyay

Languages are grouped into families that share common linguistic traits. While this approach has been successful in understanding genetic relations between diverse languages, more analyses are needed to accurately quantify their…

计算与语言 · 计算机科学 2024-10-04 Juan De Gregorio , Raúl Toral , David Sánchez

The normalized edit distance is one of the distances derived from the edit distance. It is useful in some applications because it takes into account the lengths of the two strings compared. The normalized edit distance is not defined in…

神经与进化计算 · 计算机科学 2013-12-09 Muhammad Marwan Muhammad Fuad

In this paper we examine the usefulness of two classes of algorithms Distance Methods, Discrete Character Methods (Felsenstein and Felsenstein 2003) widely used in genetics, for predicting the family relationships among a set of related…

计算与语言 · 计算机科学 2014-01-06 Taraka Rama , Sudheer Kolachina , Lakshmi Bai B

The Levenshtein distance is an important tool for the comparison of symbolic sequences, with many appearances in genome research, linguistics and other areas. For efficient applications, an approximation by a distance of smaller…

定量方法 · 定量生物学 2007-05-23 Michael Baake , Uwe Grimm , Robert Giegerich

The Swadesh approach for determining the temporal separation between two languages relies on the stochastic process of words replacement (when a complete new word emerges to represent a given concept). It is well known that the basic…

计算与语言 · 计算机科学 2025-10-28 Maurizio Serva

It is known that humans can easily read words where the letters have been jumbled in a certain way. This paper examines this problem by associating a distance measure with the jumbling process. Modifications to text were generated according…

信息检索 · 计算机科学 2011-01-05 Venkata Ravinder Paruchuri

This paper addresses the problem of deriving distance measures between parent and daughter languages with specific relevance to historical Chinese phonology. The diachronic relationship between the languages is modelled as a Probabilistic…

cmp-lg · 计算机科学 2008-02-03 Anand Raman , John Newman , Jon Patrick

In this paper, we present an algorithm for evaluating lexical similarity between a given language and several reference language clusters. As an input, we have a list of concepts and the corresponding translations in all considered…

计算与语言 · 计算机科学 2025-04-10 Karol Mikula , Mariana Sarkociová Remešíková

There is a great deal of work in cognitive psychology, linguistics, and computer science, about using word (or phrase) frequencies in context in text corpora to develop measures for word similarity or word association, going back to at…

计算与语言 · 计算机科学 2009-05-26 Rudi L. Cilibrasi , Paul M. B. Vitanyi

This work proposes a tentative model for the calculation of dimensionless distances between phonemes; sounds are described with binary distinctive features and distances show linear consistency in terms of such features. The model can be…

计算与语言 · 计算机科学 2016-11-03 Tiago Tresoldi

Deterministic automata have been traditionally studied through the point of view of language equivalence, but another perspective is given by the canonical notion of shortest-distinguishing-word distance quantifying the of states.…

计算机科学中的逻辑 · 计算机科学 2024-04-23 Wojciech Różowski

Human history leaves fingerprints in human languages. Little is known over language evolution and its study is of great importance. Here, we construct a simple stochastic model and compare its results to statistical data of real languages.…

物理与社会 · 物理学 2015-05-13 V. Schwämmle , P. M. C. de Oliveira
‹ 上一页 1 2 3 10 下一页 ›