English
Related papers

Related papers: Threshold model of language competition including …

200 papers

Bridging the performance gap between high- and low-resource languages has been the focus of much previous work. Typological features from databases such as the World Atlas of Language Structures (WALS) are a prime candidate for this, as…

Computation and Language · Computer Science 2021-01-29 Johannes Bjerva , Isabelle Augenstein

In this investigation, an optimal control problem for a stochastic mathematical model of language competition is studied. We have considered the stochastic model of language competition by adding the stochastic terms to the deterministic…

Optimization and Control · Mathematics 2020-01-03 Sakine Esmaili , M. R. Eslahchi

Language model is one of the most important modules in statistical machine translation and currently the word-based language model dominants this community. However, many translation models (e.g. phrase-based models) generate the target…

Computation and Language · Computer Science 2015-02-06 Jiajun Zhang , Shujie Liu , Mu Li , Ming Zhou , Chengqing Zong

The field of cross-lingual sentence embeddings has recently experienced significant advancements, but research concerning low-resource languages has lagged due to the scarcity of parallel corpora. This paper shows that cross-lingual word…

Computation and Language · Computer Science 2024-04-04 Zhongtao Miao , Qiyu Wu , Kaiyan Zhao , Zilong Wu , Yoshimasa Tsuruoka

The evaluation of cross-lingual semantic search models is often limited to existing datasets from tasks such as information retrieval and semantic textual similarity. We introduce Cross-Lingual Semantic Discrimination (CLSD), a lightweight…

Computation and Language · Computer Science 2025-10-10 Andrianos Michail , Simon Clematide , Rico Sennrich

Cross-lingual semantic textual similarity systems estimate the degree of the meaning similarity between two sentences, each in a different language. State-of-the-art algorithms usually employ machine translation and combine vast amount of…

Computation and Language · Computer Science 2018-07-12 Tomáš Brychcín

Bilingual word embeddings have been widely used to capture the similarity of lexical semantics in different human languages. However, many applications, such as cross-lingual semantic search and question answering, can be largely benefited…

Computation and Language · Computer Science 2019-09-10 Muhao Chen , Yingtao Tian , Haochen Chen , Kai-Wei Chang , Steven Skiena , Carlo Zaniolo

The processes leading to change in languages are manifold. In order to reduce ambiguity in the transmission of information, agreement on a set of conventions for recurring problems is favored. In addition to that, speakers tend to use…

Physics and Society · Physics 2015-06-17 Cristina-Maria Pop , Erwin Frey

We investigate learning collections of languages from texts by an inductive inference machine with access to the current datum and a bounded memory in form of states. Such a bounded memory states (BMS) learner is considered successful in…

Formal Languages and Automata Theory · Computer Science 2021-06-18 Timo Kötzing , Karen Seidel

Language and cultural diversity is a fundamental aspect of the present world. We study three modern multilingual societies -- the Basque Country, Ireland and Wales -- which are endowed with two, linguistically distant, official languages:…

Econometrics · Economics 2019-09-02 Stefan Sperlich , Jose-Ramon Uriarte

Language models generally produce grammatical text, but they are more likely to make errors in certain contexts. Drawing on paradigms from psycholinguistics, we carry out a fine-grained analysis of those errors in different syntactic…

Computation and Language · Computer Science 2025-10-30 James A. Michaelov , Catherine Arnett

Fine-tuning multilingual foundation models on specific languages often induces catastrophic forgetting, degrading performance on languages unseen in fine-tuning. While this phenomenon is widely-documented, the literature presents fragmented…

Computation and Language · Computer Science 2025-10-23 Danni Liu , Jan Niehues

Pretrained multilingual language models can help bridge the digital language divide, enabling high-quality NLP models for lower resourced languages. Studies of multilingual models have so far focused on performance, consistency, and…

Computation and Language · Computer Science 2022-10-12 Laura Cabello Piqueras , Anders Søgaard

Preference optimization techniques have become a standard final stage for training state-of-art large language models (LLMs). However, despite widespread adoption, the vast majority of work to-date has focused on first-class citizen…

Computation and Language · Computer Science 2024-07-04 John Dang , Arash Ahmadian , Kelly Marchisio , Julia Kreutzer , Ahmet Üstün , Sara Hooker

Linguistic coordination is a well-established phenomenon in spoken conversations and often associated with positive social behaviors and outcomes. While there have been many attempts to measure lexical coordination or entrainment in…

Computation and Language · Computer Science 2019-04-15 Md Nasir , Sandeep Nallan Chakravarthula , Brian Baucom , David C. Atkins , Panayiotis Georgiou , Shrikanth Narayanan

Recent developments in machine translation and multilingual text generation have led researchers to adopt trained metrics such as COMET or BLEURT, which treat evaluation as a regression problem and use representations from multilingual…

Computation and Language · Computer Science 2021-10-14 Amy Pu , Hyung Won Chung , Ankur P. Parikh , Sebastian Gehrmann , Thibault Sellam

In multilingual pretraining, the test loss of a pretrained model is heavily influenced by the proportion of each language in the pretraining data, namely the \textit{language mixture ratios}. Multilingual scaling laws can predict the test…

Computation and Language · Computer Science 2026-05-29 Xuyang Cao , Qianying Liu , Chuan Xiao , Yusuke Oda , Jiayi Wang , Pontus Stenetorp , Daisuke Kawahara , Makoto Onizuka , Sadao Kurohashi , Shuyuan Zheng

Parallel texts (bitexts) have properties that distinguish them from other kinds of parallel data. First, most words translate to only one other word. Second, bitext correspondence is noisy. This article presents methods for biasing…

cmp-lg · Computer Science 2007-05-23 I. Dan Melamed

Social mobilization often fails not for a lack of collective interest, but because of fierce competition between rival movements for the same limited pool of participants. We generalize the classic threshold model of collective behavior to…

Physics and Society · Physics 2026-05-14 Bianca Y. S. Ishikawa , José F. Fontanari

The influence of an external random field on the competition process in a nonlinear open spatially extended system is analyzed numerically. A three-component model is chosen as the competition model in which a "weak" species can move in…

Pattern Formation and Solitons · Physics 2015-12-02 S. E. Kurushina , V. V. Maximov , E. A. Shapovalova , Yu. M. Romanovskii , I. P. Zavershinskii , D. S. Garipov