中文
相关论文

相关论文: Misogyny classification of German newspaper forum …

200 篇论文

This study presents a multi-stage classification framework for detecting human values in noisy Russian language social media, validated on a random sample of 7.5 million public text posts. Drawing on Schwartz's theory of basic human values,…

计算与语言 · 计算机科学 2026-03-20 Maria Milkova , Maksim Rudnev

Hate speech is a widespread and harmful form of online discourse, encompassing slurs and defamatory posts that can have serious social, psychological, and sometimes physical impacts on targeted individuals and communities. As social media…

机器学习 · 计算机科学 2025-08-08 Santosh Chapagain , Shah Muhammad Hamdi , Soukaina Filali Boubrahimi

Crime reporting is a prevalent form of journalism with the power to shape public perceptions and social policies. How does the language of these reports act on readers? We seek to address this question with the SuspectGuilt Corpus of…

计算与语言 · 计算机科学 2020-10-16 Elisa Kreiss , Zijian Wang , Christopher Potts

The shift of public debate to the digital sphere has been accompanied by a rise in online hate speech. While many promising approaches for hate speech classification have been proposed, studies often focus only on a single language, usually…

计算与语言 · 计算机科学 2023-01-03 Ana Kotarcic , Dominik Hangartner , Fabrizio Gilardi , Selina Kurer , Karsten Donnay

Wikipedia is a community-created online encyclopedia; arguably, it is the most popular and largest knowledge resource on the Internet. Thus, reliability and neutrality are of high importance for Wikipedia. Previous research [3] reveals…

计算机与社会 · 计算机科学 2017-02-06 Olga Zagovora

In this report, we present a study of eight corpora of online hate speech, by demonstrating the NLP techniques that we used to collect and analyze the jihadist, extremist, racist, and sexist content. Analysis of the multilingual corpora…

计算与语言 · 计算机科学 2018-09-12 Tom De Smedt , Sylvia Jaki , Eduan Kotzé , Leïla Saoud , Maja Gwóźdź , Guy De Pauw , Walter Daelemans

Pretrained language models are publicly available and constantly finetuned for various real-life applications. As they become capable of grasping complex contextual information, harmful biases are likely increasingly intertwined with those…

计算与语言 · 计算机科学 2023-06-28 Sophie Jentzsch , Cigdem Turan

Migration has been a core topic in German political debate, from the postwar displacement of millions of expellees to labor migration and recent refugee movements. Studying political speech across such wide-ranging phenomena in depth has…

计算与语言 · 计算机科学 2026-04-06 Aida Kostikova , Ole Pütz , Steffen Eger , Olga Sabelfeld , Benjamin Paassen

In this study, we investigate the use of a large language model to assist in the evaluation of the reliability of the vast number of existing online news publishers, addressing the impracticality of relying solely on human expert annotators…

社会与信息网络 · 计算机科学 2025-02-14 Manuel Pratelli , John Bianchi , Fabio Pinelli , Marinella Petrocchi

Over the years, the number of users of social media has increased drastically. People frequently share their thoughts through social platforms, and this leads to an increase in hate content. In this virtual community, individuals share…

计算与语言 · 计算机科学 2024-10-21 Tonmoy Roy , Md Robiul Islam , Asif Ahammad Miazee , Anika Antara , Al Amin , Sunjim Hossain

Content moderation research has recently made significant advances, but remains limited in serving the majority of the world's languages due to the lack of resources, leaving millions of vulnerable users to online hostility. This work…

计算与语言 · 计算机科学 2025-10-28 Fitsum Gaim , Hoyun Song , Huije Lee , Changgeon Ko , Eui Jun Hwang , Jong C. Park

Women are influential online, especially in image-based social media such as Twitter and Instagram. However, many in the network environment contain gender discrimination and aggressive information, which magnify gender stereotypes and…

计算与语言 · 计算机科学 2022-04-21 Da Li , Ming Yi , Yukai He

Online antisemitism is hard to quantify. How can it be measured in rapidly growing and diversifying platforms? Are the numbers of antisemitic messages rising proportionally to other content or is it the case that the share of antisemitic…

计算机与社会 · 计算机科学 2019-10-07 Gunther Jikeli , Damir Cavar , Daniel Miehling

Student reviews often make reference to professors' physical appearances. Until recently RateMyProfessors.com, the website of this study's focus, used a design feature to encourage a "hot or not" rating of college professors. In the wake of…

计算与语言 · 计算机科学 2020-10-19 Angie Waller , Kyle Gorman

Toxic comment classification models are often found biased toward identity terms which are terms characterizing a specific group of people such as "Muslim" and "black". Such bias is commonly reflected in false-positive predictions, i.e.…

计算与语言 · 计算机科学 2022-10-18 Zhixue Zhao , Ziqi Zhang , Frank Hopfgartner

Online news outlets are grappling with the moderation of user-generated content within their comment section. We present a recommender system based on ranking class probabilities to support and empower the moderator in choosing featured…

信息检索 · 计算机科学 2023-07-17 Cedric Waterschoot , Antal van den Bosch

Understanding the impact of digital platforms on user behavior presents foundational challenges, including issues related to polarization, misinformation dynamics, and variation in news consumption. Comparative analyses across platforms and…

The ability to quantify incivility online, in news and in congressional debates, is of great interest to political scientists. Computational tools for detecting online incivility for English are now fairly accessible and potentially could…

计算与语言 · 计算机科学 2021-02-09 Anushree Hede , Oshin Agarwal , Linda Lu , Diana C. Mutz , Ani Nenkova

Forced labour is the most common type of modern slavery, and it is increasingly gaining the attention of the research and social community. Recent studies suggest that artificial intelligence (AI) holds immense potential for augmenting…

计算与语言 · 计算机科学 2022-05-06 Erick Mendez Guzman , Viktor Schlegel , Riza Batista-Navarro

Experimenting with a dataset of approximately 1.6M user comments from a Greek news sports portal, we explore how a state of the art RNN-based moderation method can be improved by adding user embeddings, user type embeddings, user biases, or…

计算与语言 · 计算机科学 2017-08-15 John Pavlopoulos , Prodromos Malakasiotis , Juli Bakagianni , Ion Androutsopoulos