English
Related papers

Related papers: Alberti's letter counts

200 papers

The availability of quantitative text analysis methods has provided new ways of analyzing literature in a manner that was not available in the pre-information era. Here we apply comprehensive machine learning analysis to the work of William…

Computation and Language · Computer Science 2024-02-14 Charles Swisher , Lior Shamir

This paper presents a new database collected from a bilingual speakers set (49), in two different languages: Spanish and Catalan. Phonetically there are significative differences between both languages. These differences have let us to…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-07 Antonio Satue-Villar , Marcos Faundez-Zanuy

Traditional linguistic theories have largely regard language as a formal system composed of rigid rules. However, their failures in processing real language, the recent successes in statistical natural language processing, and the findings…

Computation and Language · Computer Science 2020-12-02 Shuiyuan Yu , Chunshan Xu , Haitao Liu

A quantitative method is suggested, where meanings of words, and grammatic rules about these, of a vocabulary are represented by real numbers. People meet randomly, and average their vocabularies if they are equal; otherwise they either…

Physics and Society · Physics 2009-11-13 Caglar Tuncay

We collect nine corpora of representative Chinese poetry for the time span of 1046 BCE and 1644 CE for studying the history of Chinese words, collocations, and patterns. By flexibly integrating our own tools, we are able to provide new…

Computation and Language · Computer Science 2017-09-19 Chao-Lin Liu

In his pioneering research, G. K. Zipf formulated a couple of statistical laws on the relationship between the frequency of a word with its number of meanings: the law of meaning distribution, relating the frequency of a word and its…

Computation and Language · Computer Science 2022-01-19 Neus Català , Jaume Baixeries , Ramon Ferrer-Cancho , Lluís Padró , Antoni Hernández-Fernández

Recent automatic lyrics transcription (ALT) approaches focus on building stronger acoustic models or in-domain language models, while the pronunciation aspect is seldom touched upon. This paper applies a novel computational analysis on the…

Information Retrieval · Computer Science 2021-06-22 Emir Demirel , Sven Ahlback , Simon Dixon

We study the entropy of Chinese and English texts, based on characters in case of Chinese texts and based on words for both languages. Significant differences are found between the languages and between different personal styles of debating…

Computation and Language · Computer Science 2017-01-17 R. R. Xie , W. B. Deng , D. J. Wang , L. P. Csernai

For any particularly interesting theorem one proof is never enough. Instead, the first proof sets the challenge to find a more elegant method that illuminates subtle features of the math, is simpler to understand, or even avoids using…

History and Overview · Mathematics 2014-01-23 Christina Knapp , Cesar E. Silva

Regular sound correspondences constitute the principal evidence in historical language comparison. Despite the heuristic focus on regularity, it is often more an intuitive judgement than a quantified evaluation, and irregularity is more…

Computation and Language · Computer Science 2026-02-03 Frederic Blum , Johann-Mattis List

Has the style of scientific communication changed due to the growing use of large language models in the writing process? We address this question in the domain of Natural Language Processing by leveraging two data resources we create: a…

Computation and Language · Computer Science 2026-05-20 Filip Miletić , Neele Falk

Large language models (LLMs) struggle on simple tasks such as counting the number of occurrences of a letter in a word. In this paper, we investigate if ChatGPT can learn to count letters and propose an efficient solution.

Computation and Language · Computer Science 2025-02-25 Javier Conde , Gonzalo Martínez , Pedro Reviriego , Zhen Gao , Shanshan Liu , Fabrizio Lombardi

The voice laboratory permits to study the human voices using a method that is objective and noninvasive. In this work, we have studied the parameters of the human voice such as pitch, formant, jitter, shimmer and harmonic-noise ratio of a…

Neurons and Cognition · Quantitative Biology 2015-08-26 E. V. Bonzi , G. B. Grad , A. M. Maggi , M. R. Muñóz

Recent advances in large language models have created new opportunities for stylometry, the study of writing styles and authorship. Two challenges, however, remain central: training generative models when no paired data exist, and…

Computation and Language · Computer Science 2025-11-26 Mosab Rezaei , Mina Rajaei Moghadam , Abdul Rahman Shaikh , Hamed Alhoori , Reva Freedman

In this paper we study, analyse and comment rhetorical figures present in some of most interesting poetry of the first half of the twentieth century. These figures are at first traced back to some famous poet of the past and then compared…

Computation and Language · Computer Science 2018-04-03 Rodolfo Delmonte

The limited range in its abscissa of ranked letter frequency distributions causes multiple functions to fit the observed distribution reasonably well. In order to critically compare various functions, we apply the statistical model…

Computation and Language · Computer Science 2012-05-07 Wentian Li , Pedro Miramontes

Based on one million arXiv papers submitted from May 2018 to January 2024, we assess the textual density of ChatGPT's writing style in their abstracts through a statistical analysis of word frequency changes. Our model is calibrated and…

Computation and Language · Computer Science 2024-11-11 Mingmeng Geng , Roberto Trotta

Recent advances in cultural analytics and large-scale computational studies of art, literature and film often show that long-term change in the features of artistic works happens gradually. These findings suggest that conservative forces…

Computation and Language · Computer Science 2022-05-04 Artjoms Šeļa , Petr Plecháč , Alie Lassche

A universal First-Letter Law (FLL) is derived and described. It predicts the percentages of first letters for words in novels. The FLL is akin to Benford's law (BL) of first digits, which predicts the percentages of first digits in a data…

Computation and Language · Computer Science 2018-08-21 Xiaoyong Yan , Seong-Gyu Yang , Beom Jun Kim , Petter Minnhagen

This study explores the use of large language models (LLMs) to enhance datasets and improve irony detection in 19th-century Latin American newspapers. Two strategies were employed to evaluate the efficacy of BERT and GPT-4o models in…

Computation and Language · Computer Science 2025-03-31 Kevin Cohen , Laura Manrique-Gómez , Rubén Manrique
‹ Prev 1 3 4 5 6 7 10 Next ›