中文
相关论文

相关论文: Neutral evolution and turnover over centuries of E…

200 篇论文

Of basic interest is the quantification of the long term growth of a language's lexicon as it develops to more completely cover both a culture's communication requirements and knowledge space. Here, we explore the usage dynamics of words in…

计算与语言 · 计算机科学 2017-03-27 Eitan Adam Pechenick , Christopher M. Danforth , Peter Sheridan Dodds

The availability of large diachronic corpora has provided the impetus for a growing body of quantitative research on language evolution and meaning change. The central quantities in this research are token frequencies of linguistic elements…

计算与语言 · 计算机科学 2020-06-17 Andres Karjus , Richard A. Blythe , Simon Kirby , Kenny Smith

We propose a stochastic model for the number of different words in a given database which incorporates the dependence on the database size and historical changes. The main feature of our model is the existence of two different classes of…

物理与社会 · 物理学 2013-05-16 Martin Gerlach , Eduardo G. Altmann

The evolution of vocabulary in academic publishing is characterized via keyword frequencies recorded the ISI Web of Science citations database. In four distinct case-studies, evolutionary analysis of keyword frequency change through time is…

物理与社会 · 物理学 2009-02-18 R. Alexander Bentley

The neutral theory of genetic and linguistic evolution holds that the relative frequencies of variants evolve by random drift. Neutral evolution remains a plausible null model of language change. In this paper we provide evidence against…

物理与社会 · 物理学 2020-05-18 James Burridge , Tamsin Blaxter

The word-stock of a language is a complex dynamical system in which words can be created, evolve, and become extinct. Even more dynamic are the short-term fluctuations in word usage by individuals in a population. Building on the recent…

物理与社会 · 物理学 2013-04-09 Eduardo G. Altmann , Zakary L. Whichard , Adilson E. Motter

We review the task of aligning simple models for language dynamics with relevant empirical data, motivated by the fact that this is rarely attempted in practice despite an abundance of abstract models. We propose that one way to meet this…

物理与社会 · 物理学 2015-05-26 R. A. Blythe

Neural language models learn, to varying degrees of accuracy, the grammatical properties of natural languages. In this work, we investigate whether there are systematic sources of variation in the language models' accuracy. Focusing on…

计算与语言 · 计算机科学 2020-10-28 Charles Yu , Ryan Sie , Nico Tedeschi , Leon Bergen

Neutral evolution assumes that there are no selective forces distinguishing different variants in a population. Despite this striking assumption, many recent studies have sought to assess whether neutrality can provide a good description of…

种群与进化 · 定量生物学 2017-04-26 James P. O'Dwyer , Anne Kandler

Language change is a cultural evolutionary process in which variants of linguistic variables change in frequency through processes analogous to mutation, selection and genetic drift. In this work, we apply a recently-introduced method to…

计算与语言 · 计算机科学 2023-08-22 Juan Guerrero Montero , Andres Karjus , Kenny Smith , Richard A. Blythe

We introduce a new dynamic vocabulary for language models. It can involve arbitrary text spans during generation. These text spans act as basic generation bricks, akin to tokens in the traditional static vocabularies. We show that, the…

计算与语言 · 计算机科学 2024-10-14 Yanting Liu , Tao Ji , Changzhi Sun , Yuanbin Wu , Xiaoling Wang

In this paper we provide a quantitative framework for the study of phonological networks (PNs) for the English language by carrying out principled comparisons to null models, either based on site percolation, randomization techniques, or…

计算与语言 · 计算机科学 2015-06-11 Massimo Stella , Markus Brede

This discussion paper reflects on how quantitative approaches to historical linguistics interact with dataset properties. Drawing on two worked examples, we examine English data using quad-based concept modelling of Early Modern English…

计算与语言 · 计算机科学 2026-05-05 Catherine Wong , Bach Phan-Tat , Susan Fitzmaurice

How well do language models deal with quantification? In this study, we focus on 'few'-type quantifiers, as in 'few children like toys', which might pose a particular challenge for language models because the sentence components with out…

计算与语言 · 计算机科学 2023-05-29 James A. Michaelov , Benjamin K. Bergen

It is generally believed that, when a linguistic item acquires a new meaning, its overall frequency of use in the language rises with time with an S-shaped growth curve. Yet, this claim has only been supported by a limited number of case…

物理与社会 · 物理学 2017-12-04 Quentin Feltgen , Benjamin Fagard , Jean-Pierre Nadal

The availability of large linguistic data sets enables data-driven approaches to study linguistic change. The Google Books corpus unigram frequency data set is used to investigate the word rank dynamics in eight languages. We observed the…

计算与语言 · 计算机科学 2022-02-15 Alex John Quijano , Rick Dale , Suzanne Sindi

Neutral dynamics, where taxa are assumed to be demographically equivalent and their abundance is governed solely by the stochasticity of the underlying birth-death process, has proved itself as an important minimal model that accounts for…

种群与进化 · 定量生物学 2015-09-09 David Kessler , Samir Suweis , Marco Formentin , Nadav M. Shnerb

It is tempting to treat frequency trends from the Google Books data sets as indicators of the "true" popularity of various words and phrases. Doing so allows us to draw quantitatively strong conclusions about the evolution of cultural…

物理与社会 · 物理学 2020-05-28 Eitan Adam Pechenick , Christopher M. Danforth , Peter Sheridan Dodds

It is shown that a real novel shares many characteristic features with a null model in which the words are randomly distributed throughout the text. Such a common feature is a certain translational invariance of the text. Another is that…

计算与语言 · 计算机科学 2009-10-14 Sebastian Bernhardsson , Luis Enrique Correa da Rocha , Petter Minnhagen

It is now a common practice to compare models of human language processing by predicting participant reactions (such as reading times) to corpora consisting of rich naturalistic linguistic materials. However, many of the corpora used in…

‹ 上一页 1 2 3 10 下一页 ›