English
Related papers

Related papers: A Web Scraping Methodology for Bypassing Twitter A…

200 papers

Research into influence campaigns on Twitter has mostly relied on identifying malicious activities from tweets obtained via public APIs. These APIs provide access to public tweets that have not been deleted. However, bad actors can delete…

Social and Information Networks · Computer Science 2022-03-30 Christopher Torres-Lugo , Manita Pote , Alexander Nwala , Filippo Menczer

Social network has gained remarkable attention in the last decade. Accessing social network sites such as Twitter, Facebook LinkedIn and Google+ through the internet and the web 2.0 technologies has become more affordable. People are…

Social and Information Networks · Computer Science 2023-06-22 Mariam Adedoyin-Olowe , Mohamed Medhat Gaber , Frederic Stahl

This project addresses the problem of sentiment analysis in twitter; that is classifying tweets according to the sentiment expressed in them: positive, negative or neutral. Twitter is an online micro-blogging and social-networking platform…

Computation and Language · Computer Science 2015-09-15 Afroze Ibrahim Baqapuri

In contrast to much previous work that has focused on location classification of tweets restricted to a specific country, here we undertake the task in a broader context by classifying global tweets at the country level, which is so far…

Information Retrieval · Computer Science 2017-04-26 Arkaitz Zubiaga , Alex Voss , Rob Procter , Maria Liakata , Bo Wang , Adam Tsakalidis

One of the most significant current challenges in large-scale online social networks, is to establish a concise and coherent method able to collect and summarize data. Sampling the content of an Online Social Network (OSN) plays an…

Social and Information Networks · Computer Science 2015-07-07 C. A. Piña-García , Dongbing Gu

Twitter has provided a great opportunity for public libraries to disseminate information for a variety of purposes. Twitter data have been applied in different domains such as health, politics, and history. There are thousands of public…

Computers and Society · Computer Science 2018-10-01 Amir Karami , Matthew Collins

On-line social networks have grown quickly over the last few years and nowadays many people use them frequently. Furthermore the emergence of smartphones allows to access these networks any time from any physical location. Among the social…

Social and Information Networks · Computer Science 2014-04-29 Antònia Tugores , Pere Colet

Maintaining the integrity of long-term data collection is an essential scientific practice. As a field evolves, so too will that field's measurement instruments and data storage systems, as they are invented, improved upon, and made…

Physics and Society · Physics 2020-08-31 P. S. Dodds , J. R. Minot , M. V. Arnold , T. Alshaabi , J. L. Adams , D. R. Dewhurst , A. J. Reagan , C. M. Danforth

In applications involving conversational speech, data sparsity is a limiting factor in building a better language model. We propose a simple, language-independent method to quickly harvest large amounts of data from Twitter to supplement a…

Computation and Language · Computer Science 2015-04-13 Aaron Jaech , Mari Ostendorf

The broad adoption of online social networking platforms has made it possible to study communication networks at an unprecedented scale. Digital trace data can be compiled into large data sets of online discourse. However, it is a challenge…

Social and Information Networks · Computer Science 2012-12-21 Karissa McKelvey , Fil Menczer

Efficient and reliable social bot classification is crucial for detecting information manipulation on social media. Despite rapid development, state-of-the-art bot detection models still face generalization and scalability challenges, which…

Computers and Society · Computer Science 2020-06-05 Kai-Cheng Yang , Onur Varol , Pik-Mai Hui , Filippo Menczer

A comprehensive understanding of data quality is the cornerstone of measurement studies in social media research. This paper presents in-depth measurements on the effects of Twitter data sampling across different timescales and different…

Social and Information Networks · Computer Science 2020-04-07 Siqi Wu , Marian-Andrei Rizoiu , Lexing Xie

Social spam produces a great amount of noise on social media services such as Twitter, which reduces the signal-to-noise ratio that both end users and data mining applications observe. Existing techniques on social spam detection have…

Information Retrieval · Computer Science 2015-03-26 Bo Wang , Arkaitz Zubiaga , Maria Liakata , Rob Procter

The problem of clustering content in social media has pervasive applications, including the identification of discussion topics, event detection, and content recommendation. Here we describe a streaming framework for online detection and…

Social and Information Networks · Computer Science 2017-03-07 Mohsen JafariAsbagh , Emilio Ferrara , Onur Varol , Filippo Menczer , Alessandro Flammini

In a separate study, we were interested in understanding people's Q&A habits on Twitter. Finding questions within Twitter turned out to be a difficult challenge, so we considered applying some traditional NLP approaches to the problem. On…

Computation and Language · Computer Science 2020-06-16 Kyle Dent , Sharoda Paul

Streams of user-generated content in social media exhibit patterns of collective attention across diverse topics, with temporal structures determined both by exogenous factors and endogenous factors. Teasing apart different topics and…

Physics and Society · Physics 2014-03-07 A. Panisson , L. Gauvin , M. Quaggiotto , C. Cattuto

Social media platforms host discussions about a wide variety of topics that arise everyday. Making sense of all the content and organising it into categories is an arduous task. A common way to deal with this issue is relying on topic…

Computation and Language · Computer Science 2022-09-21 Dimosthenis Antypas , Asahi Ushio , Jose Camacho-Collados , Leonardo Neves , Vítor Silva , Francesco Barbieri

Since the length of microblog texts, such as tweets, is strictly limited to 140 characters, traditional Information Retrieval techniques suffer from the vocabulary mismatch problem severely and cannot yield good performance in the context…

Information Retrieval · Computer Science 2015-03-16 Runwei Qiang , Feifan Fan , Chao Lv , Jianwu Yang

Often, due to prohibitively large size or to limits to data collecting APIs, it is not possible to work with a complete network dataset and sampling is required. A type of sampling which is consistent with Twitter API restrictions is…

Social and Information Networks · Computer Science 2023-06-27 Naomi A. Arnold , Raul J. Mondragon , Richard G. Clegg

Twitter is a well-known microblogging social site where users express their views and opinions in real-time. As a result, tweets tend to contain valuable information. With the advancements of deep learning in the domain of natural language…

Computation and Language · Computer Science 2020-10-22 Mohiuddin Md Abdul Qudar , Vijay Mago