中文
相关论文

相关论文: A Crowd-Annotated Spanish Corpus for Humor Analysi…

200 篇论文

Understanding how political attention is divided and over what subjects is crucial for research on areas such as agenda setting, framing, and political rhetoric. Existing methods for measuring attention, such as manual labeling according to…

社会与信息网络 · 计算机科学 2019-09-19 Libby Hemphill , Angela M. Schöpke-Gonzalez

Social media plays an increasing role in our communication with friends and family, and our consumption of information and entertainment. Hence, to design effective ranking functions for posts on social media, it would be useful to predict…

机器学习 · 计算机科学 2022-07-13 Jane Dwivedi-Yu , Alon Y. Halevy

Humor recognition has been extensively studied with different methods in the past years. However, existing studies on humor recognition do not understand the mechanisms that generate humor. In this paper, inspired by the incongruity theory,…

计算与语言 · 计算机科学 2023-02-09 Yang Liu , Yuexian Hou

This paper describes the development of a multilingual, manually annotated dataset for three under-resourced Dravidian languages generated from social media comments. The dataset was annotated for sentiment analysis and offensive language…

Online hate speech is associated with substantial social harms, yet it remains unclear how consistently platforms enforce hate speech policies or whether enforcement is feasible at scale. We address these questions through a global audit of…

We perform a statistical analysis of emotionally annotated comments in two large online datasets, examining chains of consecutive posts in the discussions. Using comparisons with randomised data we show that there is a high level of…

Nowcasting based on social media text promises to provide unobtrusive and near real-time predictions of community-level outcomes. These outcomes are typically regarding people, but the data is often aggregated without regard to users in the…

社会与信息网络 · 计算机科学 2018-08-30 Salvatore Giorgi , Daniel Preotiuc-Pietro , Anneke Buffone , Daniel Rieman , Lyle H. Ungar , H. Andrew Schwartz

Understanding various humour styles is essential for comprehending the multifaceted nature of humour and its impact on fields such as psychology and artificial intelligence. This understanding has revealed that humour, depending on the…

计算与语言 · 计算机科学 2024-02-06 Mary Ogbuka Kenneth , Foaad Khosmood , Abbas Edalat

I present a tool which tells the quality of document or its usefulness based on annotations. Annotation may include comments, notes, observation, highlights, underline, explanation, question or help etc. comments are used for evaluative…

信息检索 · 计算机科学 2011-11-08 Archana Shukla

Socialbots, or non-human/algorithmic social media users, have recently been documented as competing for information dissemination and disruption on online social networks. Here we investigate the influence of socialbots in Mexican Twitter…

计算机与社会 · 计算机科学 2022-02-15 E. Velázquez , M. Yazdani , P. Suárez-Serrato

We build a novel database of around 285,000 notes from the Twitter Community Notes program to analyze the causal influence of appending contextual information to potentially misleading posts on their dissemination. Employing a difference in…

综合经济学 · 经济学 2024-04-04 Thomas Renault , David Restrepo Amariles , Aurore Troussel

Hate speech detection is a crucial task, especially on social media, where harmful content can spread quickly. Implementing machine learning models to automatically identify and address hate speech is essential for mitigating its impact and…

计算与语言 · 计算机科学 2025-08-19 Somaiyeh Dehghan , Mehmet Umut Sen , Berrin Yanikoglu

Sentiment analysis of social media data consists of attitudes, assessments, and emotions which can be considered a way human think. Understanding and classifying the large collection of documents into positive and negative aspects are a…

计算与语言 · 计算机科学 2020-07-16 Aditya Sharma , Alex Daniels

Crowd-sourcing is a cheap and popular means of creating training and evaluation datasets for machine learning, however it poses the problem of `truth inference', as individual workers cannot be wholly trusted to provide reliable…

机器学习 · 计算机科学 2019-02-26 Yuan Li , Benjamin I. P. Rubinstein , Trevor Cohn

We present a resource for the task of FrameNet semantic frame disambiguation of over 5,000 word-sentence pairs from the Wikipedia corpus. The annotations were collected using a novel crowdsourcing approach with multiple workers per sentence…

计算与语言 · 计算机科学 2020-06-15 Anca Dumitrache , Lora Aroyo , Chris Welty

Lexicon based sentiment analysis usually relies on the identification of various words to which a numerical value corresponding to sentiment can be assigned. In principle, classifiers can be obtained from these algorithms by comparison with…

计算与语言 · 计算机科学 2019-06-21 Mateus Machado , Evandro Ruiz , Kuruvilla Joseph Abraham

As a contribution to personality detection in languages other than English, we rely on distant supervision to create Personal-ITY, a novel corpus of YouTube comments in Italian, where authors are labelled with personality traits. The traits…

计算与语言 · 计算机科学 2020-11-16 Elisa Bassignana , Malvina Nissim , Viviana Patti

We perform a large-scale analysis of language diatopic variation using geotagged microblogging datasets. By collecting all Twitter messages written in Spanish over more than two years, we build a corpus from which a carefully selected list…

物理与社会 · 物理学 2014-11-20 Bruno Gonçalves , David Sánchez

We present an enrichment of the Hateval corpus of hate speech tweets (Basile et. al 2019) aimed to facilitate automated counter-narrative generation. Comparably to previous work (Chung et. al. 2019), manually written counter-narratives are…

There is a new generation of emoticons, called emojis, that is increasingly being used in mobile communications and social media. In the past two years, over ten billion emojis were used on Twitter. Emojis are Unicode graphic symbols, used…

计算与语言 · 计算机科学 2015-12-09 Petra Kralj Novak , Jasmina Smailović , Borut Sluban , Igor Mozetič