中文
相关论文

相关论文: Too good to be true? Predicting author profiles fr…

200 篇论文

The pervasiveness of abusive content on the internet can lead to severe psychological and physical harm. Significant effort in Natural Language Processing (NLP) research has been devoted to addressing this problem through abusive content…

计算与语言 · 计算机科学 2021-07-23 Svetlana Kiritchenko , Isar Nejadgholi , Kathleen C. Fraser

This paper explores the use of language models to predict 20 human traits from users' Facebook status updates. The data was collected by the myPersonality project, and includes user statuses along with their personality, gender, political…

社会与信息网络 · 计算机科学 2018-07-26 Andrew Cutler , Brian Kulis

This paper presents a computational approach to author profiling taking gender and language variety into account. We apply an ensemble system with the output of multiple linear SVM classifiers trained on character and word $n$-grams. We…

计算与语言 · 计算机科学 2017-07-04 Alina Maria Ciobanu , Marcos Zampieri , Shervin Malmasi , Liviu P. Dinu

Positive feedback via likes and awards is central to online governance, yet which attributes of users' posts elicit rewards -- and how these vary across authors and communities -- remains unclear. To examine this, we combine…

人机交互 · 计算机科学 2026-02-03 Agam Goyal , Charlotte Lambert , Eshwar Chandrasekharan

Trustfulness -- one's general tendency to have confidence in unknown people or situations -- predicts many important real-world outcomes such as mental health and likelihood to cooperate with others such as clinicians. While data-driven…

计算与语言 · 计算机科学 2019-04-17 Mohammadzaman Zamani , Anneke Buffone , H. Andrew Schwartz

The purpose of this study is to find evidence for supporting the hypothesis that language is the mirror of our thinking, our prejudices and cultural stereotypes. In this analysis, a questionnaire was administered to 537 people. The answers…

计算与语言 · 计算机科学 2020-07-15 P. Cutugno , D. Chiarella , R. Lucentini , L. Marconi , G. Morgavi

With the recent proliferation of the use of text classifications, researchers have found that there are certain unintended biases in text classification datasets. For example, texts containing some demographic identity-terms (e.g., "gay",…

计算与语言 · 计算机科学 2020-08-21 Guanhua Zhang , Bing Bai , Junqi Zhang , Kun Bai , Conghui Zhu , Tiejun Zhao

User generated text on social media often suffers from a lot of undesired characteristics including hatespeech, abusive language, insults etc. that are targeted to attack or abuse a specific group of people. Often such text is written…

计算与语言 · 计算机科学 2019-10-03 Sravan Babu Bodapati , Spandana Gella , Kasturi Bhattacharjee , Yaser Al-Onaizan

This article details the advances made to a system that uses artificial intelligence to identify alarming student responses. This system is built into our assessment platform to assess whether a student's response indicates they are a…

计算与语言 · 计算机科学 2023-05-16 Christopher M. Ormerod , Milan Patel , Harry Wang

Online social network analysis has attracted great attention with a vast number of users sharing information and availability of APIs that help to crawl online social network data. In this paper, we study the research studies that are…

社会与信息网络 · 计算机科学 2016-12-28 Tayfun Tuna , Esra Akbas , Ahmet Aksoy , Muhammed Abdullah Canbaz , Umit Karabiyik , Bilal Gonen , Ramazan Aygun

The context-dependent nature of online aggression makes annotating large collections of data extremely difficult. Previously studied datasets in abusive language detection have been insufficient in size to efficiently train deep learning…

计算与语言 · 计算机科学 2018-08-31 Younghun Lee , Seunghyun Yoon , Kyomin Jung

Abuse on the Internet represents a significant societal problem of our time. Previous research on automated abusive language detection in Twitter has shown that community-based profiling of users is a promising technique for this task.…

计算与语言 · 计算机科学 2019-04-09 Pushkar Mishra , Marco Del Tredici , Helen Yannakoudakis , Ekaterina Shutova

Abusive language detection has become an increasingly important task as a means to tackle this type of harmful content in social media. There has been a substantial body of research developing models for determining if a social media post…

计算与语言 · 计算机科学 2025-08-19 Raneem Alharthi , Rajwa Alharthi , Aiqi Jiang , Arkaitz Zubiaga

Gender stereotypes are pervasive beliefs about individuals based on their gender that play a significant role in shaping societal attitudes, behaviours, and even opportunities. Recognizing the negative implications of gender stereotypes,…

计算与语言 · 计算机科学 2024-04-19 Isar Nejadgholi , Kathleen C. Fraser , Anna Kerkhof , Svetlana Kiritchenko

Large language models produce human-like text that drive a growing number of applications. However, recent literature and, increasingly, real world observations, have demonstrated that these models can generate language that is toxic,…

Many adult content websites incorporate social networking features. Although these are popular, they raise significant challenges, including the potential for users to "catfish", i.e., to create fake profiles to deceive other users. This…

社会与信息网络 · 计算机科学 2017-06-30 Walid Magdy , Yehia Elkhatib , Gareth Tyson , Sagar Joglekar , Nishanth Sastry

It has been shown in the field of Author Profiling that texts may inadvertently reveal sensitive information about their authors, such as gender or age. This raises important privacy concerns that have been extensively addressed in the…

计算与语言 · 计算机科学 2024-12-18 Martin Borquez , Mikaela Keller , Michael Perrot , Damien Sileo

The wide use of social media sites and other digital technologies have resulted in an unprecedented availability of digital data that are being used to study human behavior across research domains. Although unsolicited opinions and…

社会与信息网络 · 计算机科学 2018-06-01 Nina Cesare , Christan Grant , Quynh Nguyen , Hedwig Lee , Elaine O. Nsoesie

Social media features substantial stylistic variation, raising new challenges for syntactic analysis of online writing. However, this variation is often aligned with author attributes such as age, gender, and geography, as well as more…

计算与语言 · 计算机科学 2018-04-23 Murali Raghu Babu Balusu , Taha Merghani , Jacob Eisenstein

This paper investigates gender bias in Large Language Model (LLM)-generated teacher evaluations in higher education setting, focusing on evaluations produced by GPT-4 across six academic subjects. By applying a comprehensive analytical…

计算与语言 · 计算机科学 2024-09-17 Yuanning Huang