中文
相关论文

相关论文: Stereotypical gender actions can be extracted from…

200 篇论文

User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, character repetitions, and emojis. This makes evaluating UGC translation challenging: what…

计算与语言 · 计算机科学 2026-05-13 Lydia Nishimwe , Benoît Sagot , Rachel Bawden

Semantic sentence embeddings are usually supervisedly built minimizing distances between pairs of embeddings of sentences labelled as semantically similar by annotators. Since big labelled datasets are rare, in particular for non-English…

计算与语言 · 计算机科学 2021-10-06 Marco Di Giovanni , Marco Brambilla

This paper is focused on the computational analysis of collective discourse, a collective behavior seen in non-expert content contributions in online social media. We collect and analyze a wide range of real-world collective discourse…

社会与信息网络 · 计算机科学 2012-04-18 Vahed Qazvinian , Dragomir R. Radev

This research draws upon cognitive psychology and information systems studies to anticipate user engagement and decision-making on digital platforms. By employing natural language processing (NLP) techniques and insights from cognitive bias…

人机交互 · 计算机科学 2023-07-28 Nimrod Dvir , Elaine Friedman , Suraj Commuri , Fan Yang , Jennifer Romano

We use structural topic modeling to examine racial bias in data collected to train models to detect hate speech and abusive language in social media posts. We augment the abusive language dataset by adding an additional feature indicating…

计算与语言 · 计算机科学 2020-05-28 Thomas Davidson , Debasmita Bhattacharya

The goal of our research is to contribute information about how useful the crowd is at anticipating stereotypes that may be biasing a data set without a researcher's knowledge. The results of the crowd's prediction can potentially be used…

人机交互 · 计算机科学 2018-01-11 Zeyuan Hu , Julia Strout

With increasing globalization and immigration, various studies have estimated that about half of the world population is bilingual. Consequently, individuals concurrently use two or more languages or dialects in casual conversational…

计算与语言 · 计算机科学 2022-11-01 Saurav K. Aryal , Howard Prioleau , Gloria Washington

Studies have shown that some Natural Language Processing (NLP) systems encode and replicate harmful biases with potential adverse ethical effects in our society. In this article, we propose an approach for identifying gender and racial…

计算与语言 · 计算机科学 2022-04-13 Sean Matthews , John Hudzina , Dawn Sepehr

The movie recommender system typically leverages user feedback to provide personalized recommendations that align with user preferences and increase business revenue. This study investigates the impact of gender stereotypes on such systems…

信息检索 · 计算机科学 2025-01-09 Falguni Roy , Yiduo Shen , Na Zhao , Xiaofeng Ding , Md. Omar Faruk

Currently, there is a surge of interest in fair Artificial Intelligence (AI) and Machine Learning (ML) research which aims to mitigate discriminatory bias in AI algorithms, e.g. along lines of gender, age, and race. While most research in…

计算机与社会 · 计算机科学 2021-07-29 Clarice Wang , Kathryn Wang , Andrew Bian , Rashidul Islam , Kamrun Naher Keya , James Foulds , Shimei Pan

Text is a vehicle to convey information that reflects the writer's linguistic style and communicative patterns. By studying these attributes, we can discover latent insights about the author and their underlying message. This article uses…

计算机与社会 · 计算机科学 2024-12-19 Deborah Gerhardt , Miriam Marcowitz-Bitton , W. Michael Schuster , Avshalom Elmalech , Omri Suissa , Moshe Mash

We propose Coactive Learning as a model of interaction between a learning system and a human user, where both have the common goal of providing results of maximum utility to the user. At each step, the system (e.g. search engine) receives a…

机器学习 · 计算机科学 2015-03-20 Pannaga Shivaswamy , Thorsten Joachims

This study delves into gender classification systems, shedding light on the interaction between social stereotypes and algorithmic determinations. Drawing on the "averageness theory," which suggests a relationship between a face's…

计算机与社会 · 计算机科学 2024-11-14 Miriam Doh , Anastasia Karagianni

Here we examine whether the personality dimension of openness to experience can be predicted from the individual google search history. By web scraping, individual text corpora (ICs) were generated from 214 participants with a mean number…

计算与语言 · 计算机科学 2024-04-02 Markus J. Hofmann , Markus T. Jansen , Christoph Wigbels , Benny Briesemeister , Arthur M. Jacobs

Humans carry stereotypic tacit assumptions (STAs) (Prince, 1978), or propositional beliefs about generic concepts. Such associations are crucial for understanding natural language. We construct a diagnostic set of word prediction prompts to…

计算与语言 · 计算机科学 2020-06-17 Nathaniel Weir , Adam Poliak , Benjamin Van Durme

Semantic representations are integral to natural language processing, psycholinguistics, and artificial intelligence. Although often derived from internet text, recent years have seen a rise in the popularity of behavior-based (e.g., free…

计算与语言 · 计算机科学 2024-12-09 Zak Hussain , Rui Mata , Ben R. Newell , Dirk U. Wulff

Generating models from large data sets -- and determining which subsets of data to mine -- is becoming increasingly automated. However choosing what data to collect in the first place requires human intuition or experience, usually supplied…

计算机与社会 · 计算机科学 2014-05-20 Josh C. Bongard , Paul D. H. Hines , Dylan Conger , Peter Hurd , Zhenyu Lu

This paper describes the analysis of quantitative characteristics of frequent sets and association rules in the posts of Twitter microblogs related to different event discussions. For the analysis, we used a theory of frequent sets,…

社会与信息网络 · 计算机科学 2013-10-15 Bohdan Pavlyshenko

Large language models exhibit societal biases associated with demographic information, including race, gender, and others. Endowing such language models with personalities based on demographic data can enable generating opinions that align…

人工智能 · 计算机科学 2024-02-29 Seungjong Sun , Eungu Lee , Dongyan Nan , Xiangying Zhao , Wonbyung Lee , Bernard J. Jansen , Jang Hyun Kim

The Stereotype Content model (SCM) states that we tend to perceive minority groups as cold, incompetent or both. In this paper we adapt existing work to demonstrate that the Stereotype Content model holds for contextualised word embeddings,…

计算与语言 · 计算机科学 2022-10-27 Eddie L. Ungless , Amy Rafferty , Hrichika Nag , Björn Ross