English
Related papers

Related papers: Forensic Authorship Analysis of Microblogging Text…

200 papers

Authorship attribution (AA), which is the task of finding the owner of a given text, is an important and widely studied research topic with many applications. Recent works have shown that deep learning methods could achieve significant…

Computation and Language · Computer Science 2021-03-23 Zhiqiang Hu , Roy Ka-Wei Lee , Lei Wang , Ee-Peng Lim , Bo Dai

In this paper we present a novel methodology for identifying scholars with a Twitter account. By combining bibliometric data from Web of Science and Twitter users identified by Altmetric.com we have obtained the largest set of individual…

Digital Libraries · Computer Science 2017-12-18 Rodrigo Costas , Jeroen van Honk , Thomas Franssen

Using machine learning algorithms, including deep learning, we studied the prediction of personal attributes from the text of tweets, such as gender, occupation, and age groups. We applied word2vec to construct word vectors, which were then…

Computers and Society · Computer Science 2017-12-27 Take Yo , Kazutoshi Sasahara

It is very critical to analyze messages shared over social networks for cyber threat intelligence and cyber-crime prevention. In this study, we propose a method that leverages both domain-specific word embeddings and task-specific features…

Computation and Language · Computer Science 2019-06-04 Semih Yagcioglu , Mehmet Saygin Seyfioglu , Begum Citamak , Batuhan Bardak , Seren Guldamlasioglu , Azmi Yuksel , Emin Islam Tatli

In this paper, we attempt to classify tweets into root categories of the Amazon browse node hierarchy using a set of tweets with browse node ID labels, a much larger set of tweets without labels, and a set of Amazon reviews. Examining…

Social and Information Networks · Computer Science 2015-11-30 Matthew Long , Aditya Jami , Ashutosh Saxena

Notwithstanding recent work which has demonstrated the potential of using Twitter messages for content-specific data mining and analysis, the depth of such analysis is inherently limited by the scarcity of data imposed by the 140 character…

Social and Information Networks · Computer Science 2016-11-17 Adham Beykikhoshk , Ognjen Arandjelovic , Dinh Phung , Svetha Venkatesh

Predicting the quality of a text document is a critical task when presented with the problem of measuring the performance of a document before its release. In this work, we evaluate various features including those extracted from the text…

Computation and Language · Computer Science 2019-10-28 Manirupa Das , Renhao Cui

The identification of spam messages on social networks is a very challenging task. Social media sites like Twitter \& Facebook attracts a lot of users and companies to advertise and attract users of personal gains. These advertisements most…

Social and Information Networks · Computer Science 2020-10-27 Prakamya Mishra

With social media datasets being increasingly shared by researchers, it also presents the caveat that those datasets are not always completely replicable. Having to adhere to requirements of platforms like Twitter, researchers cannot…

Digital Libraries · Computer Science 2018-03-08 Arkaitz Zubiaga

To analyse large numbers of texts, social science researchers are increasingly confronting the challenge of text classification. When manual labeling is not possible and researchers have to find automatized ways to classify texts, computer…

Computation and Language · Computer Science 2023-10-10 Karina Shyrokykh , Maksym Girnyk , Lisa Dellmuth

Social network and publishing platforms, such as Twitter, support the concept of verification. Verified accounts are deemed worthy of platform-wide public interest and are separately authenticated by the platform itself. There have been…

Social and Information Networks · Computer Science 2019-03-13 Indraneil Paul , Abhinav Khattar , Ponnurangam Kumaraguru , Manish Gupta , Shaan Chopra

In recent years, people spend a lot of time on social networks. They use social networks as a place to comment on personal or public events. Thus, a large amount of information is generated and shared daily in these networks. Using such a…

Social and Information Networks · Computer Science 2020-10-05 Parinaz Rahimizadeh , Mohammad Javad Shayegan

Twitter is a social media giant famous for the exchange of short, 140-character messages called "tweets". In the scientific community, the microblogging site is known for openness in sharing its data. It provides a glance into its millions…

Social and Information Networks · Computer Science 2013-06-24 Fred Morstatter , Jürgen Pfeffer , Huan Liu , Kathleen M. Carley

Text-based personality prediction by computational models is an emerging field with the potential to significantly improve on key weaknesses of survey-based personality assessment. We investigate 3848 profiles from Twitter with self-labeled…

Computation and Language · Computer Science 2021-09-15 Partha Kadambi

In recent years, numerous studies have inferred personality and other traits from people's online writing. While these studies are encouraging, more information is needed in order to use these techniques with confidence. How do linguistic…

Computation and Language · Computer Science 2015-04-27 Eben M. Haber

Nowadays, with the rise of Internet access and mobile devices around the globe, more people are using social networks for collaboration and receiving real-time information. Twitter, the microblogging that is becoming a critical source of…

Cryptography and Security · Computer Science 2020-12-02 Sepideh Bazzaz Abkenar , Mostafa Haghi Kashani , Mohammad Akbari , Ebrahim Mahdipour

Twitter is a popular social network platform where users can interact and post texts of up to 280 characters called tweets. Hashtags, hyperlinked words in tweets, have increasingly become crucial for tweet retrieval and search. Using…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-01-29 Vibhuti Gupta , Rattikorn Hewett

Text actionability detection is the problem of classifying user authored natural language text, according to whether it can be acted upon by a responding agent. In this paper, we propose a supervised learning framework for domain-aware,…

Information Retrieval · Computer Science 2015-11-04 Nemanja Spasojevic , Adithya Rao

In practice, training language models for individual authors is often expensive because of limited data resources. In such cases, Neural Network Language Models (NNLMs), generally outperform the traditional non-parametric N-gram models.…

Computation and Language · Computer Science 2016-02-18 Zhenhao Ge , Yufang Sun , Mark J. T. Smith

This paper investigates the stability of Twitter counts of scientific publications over time. For this, we conducted an analysis of the availability statuses of over 2.6 million Twitter mentions received by the 1,154 most tweeted scientific…

Digital Libraries · Computer Science 2020-02-25 Zhichao Fang , Jonathan Dudek , Rodrigo Costas