English
Related papers

Related papers: Learning about Spanish dialects through Twitter

200 papers

In recent years, social bots have been using increasingly more sophisticated, challenging detection strategies. While many approaches and features have been proposed, social bots evade detection and interact much like humans making it…

Social and Information Networks · Computer Science 2018-12-20 Isa Inuwa-Dutse , Bello Shehu Bello , Ioannis Korkontzelos

Previous studies have shown that Twitter users have biases to tweet from certain locations, locational bias, and during certain hours, temporal bias. We used three years of geolocated Twitter Data to quantify these biases and test our…

Social and Information Networks · Computer Science 2016-09-15 Aiman Soliman , Kiumars Soltani , Anand Padmanabhan , Shaowen Wang

Twitter continuously tightens the access to its data via the publicly accessible, cost-free standard APIs. This especially applies to the follow network. In light of this, we successfully modified a network sampling method to work…

Social and Information Networks · Computer Science 2021-01-13 Felix Victor Münch , Ben Thies , Cornelius Puschmann , Axel Bruns

The emergence of large stores of transactional data generated by increasing use of digital devices presents a huge opportunity for policymakers to improve their knowledge of the local environment and thus make more informed and better…

Computers and Society · Computer Science 2017-10-27 Graham McNeill , Jonathan Bright , Scott A. Hale

Given the rapidly evolving landscape of linguistic prevalence, whereby a majority of the world's existing languages are dying out in favor of the adoption of a comparatively fewer set of languages, the factors behind this phenomenon has…

Physics and Society · Physics 2021-06-30 Sayat Mimar , Mariamo Mussa Juane , Jorge Mira , Juyong Park , Alberto P. Munuzuri , Gourab Ghoshal

In applications involving conversational speech, data sparsity is a limiting factor in building a better language model. We propose a simple, language-independent method to quickly harvest large amounts of data from Twitter to supplement a…

Computation and Language · Computer Science 2015-04-13 Aaron Jaech , Mari Ostendorf

Although linguistic typology has a long history, computational approaches have only recently gained popularity. The use of distributed representations in computational linguistics has also become increasingly popular. A recent development…

Computation and Language · Computer Science 2017-11-16 Johannes Bjerva , Isabelle Augenstein

Cultural areas represent a useful concept that cross-fertilizes diverse fields in social sciences. Knowledge of how humans organize and relate their ideas and behavior within a society helps to understand their actions and attitudes towards…

Computation and Language · Computer Science 2023-04-19 Thomas Louf , Bruno Gonçalves , Jose J. Ramasco , David Sanchez , Jack Grieve

Although the prediction of dialects is an important language processing task, with a wide range of applications, existing work is largely limited to coarse-grained varieties. Inspired by geolocation research, we propose the novel task of…

Computation and Language · Computer Science 2020-12-08 Muhammad Abdul-Mageed , Chiyu Zhang , AbdelRahim Elmadany , Lyle Ungar

Twitter has rapidly emerged as one of the largest worldwide venues for written communication. Thanks to the ease with which vast quantities of tweets can be mined, Twitter has also become a source for studying modern linguistic style. The…

Social and Information Networks · Computer Science 2014-01-24 James R. A. Davenport , Robert DeLine

Machine learning techniques have conquered many different tasks in speech and natural language processing, such as speech recognition, information extraction, text and speech generation, and human machine interaction using natural language…

Computation and Language · Computer Science 2025-03-18 Sebastian Möller , Pia Knoeferle , Britta Schulte , Nils Feldhus

Twitter is a well-known microblogging social site where users express their views and opinions in real-time. As a result, tweets tend to contain valuable information. With the advancements of deep learning in the domain of natural language…

Computation and Language · Computer Science 2020-10-22 Mohiuddin Md Abdul Qudar , Vijay Mago

Artificial intelligence approaches are being adapted to many research areas, including digital humanities. We built a methodology for large-scale analyses in folkloristics. Using machine learning and natural language processing, we…

Computation and Language · Computer Science 2025-10-22 Tjaša Arčon , Marko Robnik-Šikonja , Polona Tratnik

This study explores how different weather conditions influence public sentiment on social media, focusing on Twitter data from the UK. By considering climate and linguistic baselines, we improve the accuracy of weather-related sentiment…

Human-Computer Interaction · Computer Science 2024-07-11 James C. Young , Rudy Arthur , Hywel T. P. Williams

The advent of geolocated ICT technologies opens the possibility of exploring how people use space in cities, bringing an important new tool for urban scientists and planners, especially for regions where data is scarce or not available.…

Norway has a large amount of dialectal variation, as well as a general tolerance to its use in the public sphere. There are, however, few available resources to study this variation and its change over time and in more informal areas, \eg…

Computation and Language · Computer Science 2021-04-13 Jeremy Barnes , Petter Mæhlum , Samia Touileb

The use of linguistic typological resources in natural language processing has been steadily gaining more popularity. It has been observed that the use of typological information, often combined with distributed language representations,…

Computation and Language · Computer Science 2020-05-06 Alexander Gutkin , Tatiana Merkulova , Martin Jansche

In the Middle Ages texts were learned by heart and spread using oral means of communication from generation to generation. Adaptation of the art of prose and poems allowed keeping particular descriptions and compositions characteristic for…

Computation and Language · Computer Science 2021-09-03 Arianna Di Bernardo , Simone Poetto , Pietro Sillano , Beatrice Villata , Weronika Sójka , Zofia Piętka-Danilewicz , Piotr Pranke

Resources in high-resource languages have not been efficiently exploited in low-resource languages to solve language-dependent research problems. Spanish and French are considered high resource languages in which an adequate level of data…

Computation and Language · Computer Science 2023-12-13 Fatimah Alzamzami , Abdulmotaleb El Saddik

In this era of digitization, knowing the user's sociolect aspects have become essential features to build the user specific recommendation systems. These sociolect aspects could be found by mining the user's language sharing in the form of…

Computation and Language · Computer Science 2018-04-13 Barathi Ganesh HB , Anand Kumar M , Soman KP